NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-08-01M4DA8WP0XZNVAPWYEXAHZJYA
Run Details
Model:
Claude Haiku 5.5 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
3.3s
Cost:
$0.000196
Provider:
Claude Platform on AWS
Evaluation summary
Score
Evaluator
Details Link
0.00%
Matches word count
(details)
86.44%
Dialogue to Total Word Ratio
(details)
43.2204%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
"Right, the plan's simple," Molly whispered, adjusting her fake moustache. "We slip in through the back window while the clown's doing balloon animals." "Balloon animals? Molly, that clown has been doing the same poodle for three hours." "Which is why nobody will notice us. Now, Bill, the presents are stacked by the fireplace. Grab the big ones first." "What if they're just socks?" "Socks are for amateurs. Kids always get the shiny boxes." "And the bouncy castle? Do we deal with that?" "Leave it. Nobody steals a bouncy castle, Bill. It's too conspicuous." "Fair point. Oh, Molly, there's a little boy watching us from the hallway." "Tell him we're the party magicians. Quick, wave your wand." "I don't have a wand. I have a rubber chicken." "Then wave the chicken, you idiot!" "Hi there, kid. Want to see a disappearing act?" "Bill, he's holding a juice box and he's not impressed." "Nobody's impressed by a chicken, Molly."
created_at
"2026-10-08T08:32:27.347Z"
10/8/2026, 8:32:27 AM
model
"anthropic/claude-haiku-5.5-20261007"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
739
moderation_latency
(null)
generation_time
3238
tokens_prompt
75
tokens_completion
243
native_tokens_prompt
95
native_tokens_completion
372
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.0001955
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791448347-RMFjMuBZ1a3laL4G2Wvt"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791448347-8Usr2hOsxRqi0Icc7Qsi"
upstream_id
"msg_011CfpTd5BM7uxLgLnpMDpWZ"
provider_responses
0
endpoint_id
"b6cdf493-3a61-441f-a134-e7944fddb980"
id
"msg_011CfpTd5BM7uxLgLnpMDpWZ"
is_byok
false
latency
739
model_permaslug
"anthropic/claude-haiku-5.5-20261007"
provider_name
"Claude Platform on AWS"
status
200
total_cost
0.0001955
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Claude Platform on AWS"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
0.00%
Matches word count
n/a
neededClean
false
words
157
86.44%
Dialogue to Total Word Ratio
Ratio: 96.18%, Deviation: 6.18%
neededClean
false
wordsTotal
157
wordsDialogue
151
43.2204%