Back to programming-model index

MODEL EVIDENCE / OpenRouter #12

Poolside: Laguna S 2.1 (free)

The model identity, provider record, and independently verified benchmark results in one place. The number above is OpenRouter discovery order, not benchmark rank.

poolside/laguna-s-2.1
Official source verifiedComplete result
Fast numbersCaptured facts and latest status
OpenRouter feed rank
OpenRouter #12
Context
262144
Input
text
Output
text
Reasoning flag
Yes
Prompt price
$0.000 / 1M tokens
Completion price
$0.000 / 1M tokens
Latest run
2026-08-31T21:35:59.284247+00:00
View denominator
50

DEEP BENCHMARK

What the run established.

Counts stay visible for auditability. A score appears only when the Most Annoying 50-trap view satisfies the independent checks.

MOST ANNOYING BENCHMARK

47/ 50

Complete result · 47 verified mistakes · 3 clean results · 0 unresolved

Inspect run receipt
Packs & trend

Pack 1: 10 verified mistakes, 0 clean results, 0 unresolved; Pack 2: 9 verified mistakes, 1 clean results, 0 unresolved; Pack 3: 9 verified mistakes, 1 clean results, 0 unresolved; Pack 4: 10 verified mistakes, 0 clean results, 0 unresolved; Pack 5: 9 verified mistakes, 1 clean results, 0 unresolved

2026-08-28T05:52:59.444809+00:00: 46; 2026-08-31T21:35:59.284247+00:00: 47

Score history

Model runs, then future slots.

  1. 018%
  2. 026%
  3. 03NOT RUN
  4. 04NOT RUN
  5. 05NOT RUN
  6. 06NOT RUN
  7. 07NOT RUN
  8. 08NOT RUN
  9. 09NOT RUN
  10. 10NOT RUN
  11. 11NOT RUN
  12. 12NOT RUN
  13. 13NOT RUN
  14. 14NOT RUN
  15. 15NOT RUN
  16. 16NOT RUN
  17. 17NOT RUN
  18. 18NOT RUN
  19. 19NOT RUN
  20. 20NOT RUN
  21. 21NOT RUN
  22. 22NOT RUN
  23. 23NOT RUN
  24. 24NOT RUN
Verified mistakesClean resultsUnresolved / pendingMissing / no recorded result

Most Annoying verified mistake receipts

Only receipts from the Most Annoying view are listed here. Tolerable-view receipts are separate and are not added to this score.

  1. Alias TraceT001

    The model's exact answer did not match the frozen expected output.

    sha256:2a6cccb60a58a6c55e6c5e8f7f23d7b1eb34ccb925a0c4cb226958b857e9c0be
  2. Precedence TraceT002

    The model's exact answer did not match the frozen expected output.

    sha256:36753a0dfc788ee175d5982bd374e3bc908af122d3608a7bef2845bd94baea38
  3. Loop TraceT003

    The model's exact answer did not match the frozen expected output.

    sha256:5027be8970d2dc511c7bc9f4068492be2979e97ae07f58c7a75ca7a0dc0967d0
  4. Slice TraceT004

    The model's exact answer did not match the frozen expected output.

    sha256:a5702ede1b37a705bcdcb1c355e1c7cd772b81d481480f5a9722d61c82d4eaff
  5. Alias TraceT005

    The model's exact answer did not match the frozen expected output.

    sha256:fe7871738822ff848793223f7775772ba2defafc6a231e8637510c5b822c606b
  6. Precedence TraceT006

    The model's exact answer did not match the frozen expected output.

    sha256:4ef26ce9929a072a419f8b00e8c1351fd97f557520b6b38c81b91c9e44cd856d
  7. Loop TraceT007

    The model's exact answer did not match the frozen expected output.

    sha256:bea565503c83ebc05fddeb4e981cbc72b5d0fbfe289a4d19be34e2c7fe29444f
  8. Slice TraceT008

    The model's exact answer did not match the frozen expected output.

    sha256:4a1e29c7adc6145c678d241dd38b9f2a10974c76cd0f78bfacd4bbd16d9257f7
  9. Alias TraceT009

    The model's exact answer did not match the frozen expected output.

    sha256:cad5640f0de8c45390a704f11d88e923e6681c09f13b148597906c30eb8ad433
  10. Precedence TraceT010

    The model's exact answer did not match the frozen expected output.

    sha256:7b58330165c70daecbee2c01cf0555afe21e8146ed1a3c4fcd68bdac65bc063a
  11. Alias TraceT021

    The model's exact answer did not match the frozen expected output.

    sha256:5fcc8e8986f6f0fb4cfad7ad533482cef8285c683005d608377ae846267913ff
  12. Precedence TraceT022

    The model's exact answer did not match the frozen expected output.

    sha256:16a4d867ee524c8e454493f41b268b1a8cb12b0b9c259b9b779f246d046d1da9
  13. Loop TraceT023

    The model's exact answer did not match the frozen expected output.

    sha256:ad047465afe5eead5c47c04d5c84f8d9079f91a0bb4ec583b3ef98c295289d7a
  14. Slice TraceT024

    The model's exact answer did not match the frozen expected output.

    sha256:ffbcf70a76b2b2073017538b94491c3db0bd8f923a013dbd6b7eb71e1f3c4325
  15. Alias TraceT025

    The model's exact answer did not match the frozen expected output.

    sha256:ed0e9c034621ca351029694b0f971856770a9f2b19018c18c6972f6927a98260
  16. Precedence TraceT026

    The model's exact answer did not match the frozen expected output.

    sha256:b0bce13318982da496c6c2709f997e09a507e94c67335cfe67c07a19ebe74fd0
  17. Slice TraceT028

    The model's exact answer did not match the frozen expected output.

    sha256:60070b75fd34b2f8b45df4998d24dd9208335258b0ffc5d501a8670b9a317207
  18. Alias TraceT029

    The model's exact answer did not match the frozen expected output.

    sha256:972cbaf8c2568534037989faa199666390290e811f5aede8dcc505a01c80f48c
  19. Precedence TraceT030

    The model's exact answer did not match the frozen expected output.

    sha256:12c42eb972343943b8296d32ce059345fd767c12573e96b498e412e2e0a1a9ec
  20. Alias TraceT041

    The model's exact answer did not match the frozen expected output.

    sha256:29957be3349a75c5e2a29397d3feb7db5a4d945717d4b8cab09ccfa6e4208d14
  21. Precedence TraceT042

    The model's exact answer did not match the frozen expected output.

    sha256:ae14076eb043ff122f6a2257d2efbca4c6c45237d33254ef4e1e6cb1580cb6c3
  22. Loop TraceT043

    The model's exact answer did not match the frozen expected output.

    sha256:4917e13667ed0f819b883d0ea1a97c71a835357d55bb7b1903089cffec4f9fea
  23. Alias TraceT045

    The model's exact answer did not match the frozen expected output.

    sha256:a5f2a6f4771fba09d5866de1a29262fa881e19cf8a18d75e4f89646215880da4
  24. Precedence TraceT046

    The model's exact answer did not match the frozen expected output.

    sha256:6f62c449478f511a573ec90575393b6bf30d55c61e15e374eb98c780661f9dff
  25. Loop TraceT047

    The model's exact answer did not match the frozen expected output.

    sha256:ebcb4e421b40d44571acd195c86d13fcee740ed8c589fd771867a3e6a893e3f5
  26. Slice TraceT048

    The model's exact answer did not match the frozen expected output.

    sha256:90534a82995c4645662c44730618cbd8d1fbe81104677c47d7f9764a606c16d9
  27. Alias TraceT049

    The model's exact answer did not match the frozen expected output.

    sha256:77c8764695eb5b7d2194ba7d054a1464189cdc5fbf6c159fa737da97cbfc52c0
  28. Precedence TraceT050

    The model's exact answer did not match the frozen expected output.

    sha256:5ff2d5f05be2986ae403e1d3bc08bee5c441792adc4aa470402369c0df967816
  29. Alias TraceT061

    The model's exact answer did not match the frozen expected output.

    sha256:78ee8c2697b388caa97289bbc11bebfab324d592b2b0253ebe365eb6495a3747
  30. Precedence TraceT062

    The model's exact answer did not match the frozen expected output.

    sha256:c7496a6a656194bd34cb6d2b6e4790ccd6e40fb68b71b93c8d7480c869a173df
  31. Loop TraceT063

    The model's exact answer did not match the frozen expected output.

    sha256:553d6c93a5b985f353490f99827e01968c14af877c8f7ea91dbca5bdddd3d0a2
  32. Slice TraceT064

    The model's exact answer did not match the frozen expected output.

    sha256:323666d2b8a5bf3e0e03e32e9e2f67b60d4c5ca2a72a6ac1b6b0eb0a6d780710
  33. Alias TraceT065

    The model's exact answer did not match the frozen expected output.

    sha256:1a12b98becc21ac39908d9fd5bbff6bb914b4d0afe4b8e343184476a41dd0b57
  34. Precedence TraceT066

    The model's exact answer did not match the frozen expected output.

    sha256:6754d62cd7a871cf2c7e7688d21d2b2164eb8105b757ce0dbcd8eb21eb3f49ef
  35. Loop TraceT067

    The model's exact answer did not match the frozen expected output.

    sha256:f40b4be75ce4843dbfc5318549bdc1b7450985c6f524dd94d422d12173face22
  36. Slice TraceT068

    The model's exact answer did not match the frozen expected output.

    sha256:66375cc518a221dc4da5c280fbd34dde8f0c6ac2c7a5a362f65123863c413aff
  37. Alias TraceT069

    The model's exact answer did not match the frozen expected output.

    sha256:271b8a435b4195049901a07eef972c78ce341ccc37ebb108cdfc5321bbc2b449
  38. Precedence TraceT070

    The model's exact answer did not match the frozen expected output.

    sha256:18293d2d48475d21a187d779e3d5a6b50575ab922db688d6fa09ba14a1305627
  39. Alias TraceT081

    The model's exact answer did not match the frozen expected output.

    sha256:3f29e9e3f411ae272eb9bf2b26e0e6f148c12acdce3013e102abbce1d6a49aae
  40. Precedence TraceT082

    The model's exact answer did not match the frozen expected output.

    sha256:a4811503a648155fc009d304bddabfda2c65d2897ba873aa17b47e33d4cd4e98
  41. Loop TraceT083

    The model's exact answer did not match the frozen expected output.

    sha256:9ea2b968f9ae18404d48486dbc75aeb1f026d497b2d443574e1f1c3fa4e05914
  42. Alias TraceT085

    The model's exact answer did not match the frozen expected output.

    sha256:6e04bc70ff5711baee01e9b23f1e2a6a47b84b8e118dd389526107ff7c31129b
  43. Precedence TraceT086

    The model's exact answer did not match the frozen expected output.

    sha256:98be27be735f456ba36065784d9c5b1e6c38e4100fe84d82715a0a080b1dca7e
  44. Loop TraceT087

    The model's exact answer did not match the frozen expected output.

    sha256:aba6a881df8c23c57aacad7869f698e51b2fc69e21da4915a117e7744880ddf1
  45. Slice TraceT088

    The model's exact answer did not match the frozen expected output.

    sha256:c888d3a0c55e1f5c318d93e9d453a7ffb4f800edc70ee42a41d871581593c424
  46. Alias TraceT089

    The model's exact answer did not match the frozen expected output.

    sha256:ee8dc37ec278b0b1d9e4e38d613f210cfe1fe1074eff3a4cfc8a41bbb15893f3
  47. Precedence TraceT090

    The model's exact answer did not match the frozen expected output.

    sha256:34e135fd46d013c3e2949552932dee0125bd0e5332a65b0febc6f4d1bda5b655

OFFICIAL SOURCES

Provider record.

The linked publisher-controlled source identifies this model or its official model family. Provider descriptions remain separate from benchmark findings.

STATUSOfficial source verifiedVerified: 2026-08-27T00:00:00+00:00Research method: official primary-web fallback

Important: Statements in linked publisher material are provider claims. They are not findings of this benchmark.

METHOD & LIMITS

What this page can say.

Every public leaderboard score is the latest Most Annoying 50-trap view. The 100-task run receipt is the wider audit bundle, not the benchmark denominator. Benchmark rank #1 means fewest verified mistakes among complete runs; the list can still be sorted by most mistakes for readability. OpenRouter feed rank is separate discovery metadata and is kept out of the benchmark rank column. Two-week average reporting will use the last four complete 100-task cycles when enough cycles exist. Individual verified mistakes, called catches internally, require primary and shadow checks to reproduce the exact mismatch from the recorded reply against the frozen expected output. Retrieval is not signature verification; auditors should recompute receipt roots and verify signatures against the published issuer key. Routing errors, timeouts, incomplete answers, and missing evidence stay outside the ranked score. Do not judge any model solely from this benchmark. Ranked complete runs must use the same 50 traps, scoring rules, model settings, and retry policy.

  1. Dual-check reproduction

    For each listed Most Annoying mistake, primary and shadow verification must both reproduce the exact mismatch from the recorded reply against the frozen expected output. Otherwise the item remains unresolved.

  2. Certified denominator

    The rankable score is the 50-trap Most Annoying view. The run receipt may say 100 tasks because it covers the wider audit bundle; that object is evidence, not a separate leaderboard denominator.

  3. Signature and feed rank

    Independent verification means recomputing the receipt root and checking the signature against the published issuer public key. OpenRouter rank stays discovery metadata and does not affect benchmark rank.

Reproduce this result

  1. Inspect the run receipt.
  2. Inspect each Most Annoying mistake receipt.
  3. Verify each receipt root and signature.
  4. Wait for the full proof bundle if it is not released yet.
  5. Rerun the exact prompt and expected-output comparison.

Claim ceiling: OpenRouter rank and metadata describe the captured discovery feed. Provider claims are separate from this benchmark. Scores describe only the most recent public test pack and should not be used alone to judge the model generally.

Feed root
sha256:61e4106b1d08b7ef16ec9d332548214d01937a628e573fcf9a3091f601dfcb76
Model metadata root
sha256:08ab99ab323ebe0abdedae9fd2897e2c0d77a8c4cb54f1612e6914d769c265dc
Official-source registry root
sha256:e5606a5133065340a402faa6c24a6bb3fb36d82809003a75bae28e85d7ec8e68

QUOTE THE RECORD

Share the exact page, not a screenshot without context.

Discuss this result in:

GET A SECOND EXPLANATION

Ask your favorite AI what this means.

Copy a bounded prompt that includes this model page and its claim ceiling.