NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD80Q0DGMRS5B0CRTV8AXXY
Run Details
Model:
Laguna S 2.1
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
11.4s
Cost:
$0.000059
Provider:
Poolside
Evaluation summary
Score
Evaluator
Details Link
1.04%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
0.5180%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly adjusted her black ski mask and studied the party venue through binoculars. The inflatable castle bobbed in the breeze like a giant thumb. "Bill," she whispered, "are you sure this is legal?" Bill consulted his clipboard, which was covered in crayon drawings. "Technically, we're not stealing from children. We're liberating gifts from capitalist excess." Molly nodded approvingly. "And the security?" "Three parents, one diabetic grandfather, and a golden retriever named Sparkle." Bill paused. "The dog's the real threat." "Right. What's our entry strategy?" "Pizza delivery. Kids love pizza, parents expect pizza, nobody questions pizza." Bill grinned. "I'll wear the costume." Molly winced. "You weigh two hundred pounds. You cannot convincingly portray a teenage pizza delivery boy." "Says the woman who thinks we can steal from a children's birthday party without getting caught." The inflatable castle deflated slightly in the wind. Somewhere inside, children shrieked with joy. Molly checked her watch. "We execute at cake-cutting time. Maximum chaos, minimum witnesses." "Perfect," Bill said, already practicing his 'youthful' pizza delivery voice. "I'm thinking... 'Sup dudes?'"
created_at
"2026-07-25T18:18:59.968Z"
7/25/2026, 6:18:59 PM
model
"poolside/laguna-s-2.1-20260720"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
150
moderation_latency
(null)
generation_time
2500
tokens_prompt
75
tokens_completion
307
native_tokens_prompt
109
native_tokens_completion
283
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
96
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.00005886
router
(null)
provider_responses
0
endpoint_id
"ab7f9e97-1af5-4b81-b099-cc932f54234e"
id
"chatcmpl-bab22a9b95394e388b159b6ccf770d0a"
is_byok
false
latency
150
model_permaslug
"poolside/laguna-s-2.1-20260720"
provider_name
"Poolside"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785003539-UG9Kta4VRHxgxQ9Bpxuc"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785003539-FF6DfK86V7dUXrAYKQko"
upstream_id
"chatcmpl-bab22a9b95394e388b159b6ccf770d0a"
total_cost
0.00005886
cache_discount
0.00000864
upstream_inference_cost
0
provider_name
"Poolside"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
1.04%
Matches word count
n/a
neededClean
false
words
174
0.00%
Dialogue to Total Word Ratio
Ratio: 61.14%, Deviation: 51.14%
neededClean
false
wordsTotal
175
wordsDialogue
107
0.5180%