NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-08-01M4D9YS8NPZSMBW27JHVDW8VJ
Run Details
Model:
Claude Haiku 5.5 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
3.2s
Cost:
$0.000208
Provider:
Claude Platform on AWS
Evaluation summary
Score
Evaluator
Details Link
0.00%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
0.0004%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the bouncy castle, peering through the hedge at the backyard party. "Eleven kids, one clown, and a pile of presents taller than you, Bill." Bill adjusted his fake mustache. "Good. Clowns are easy. Remind me the plan?" "We slip in as the magician's assistants. Nobody checks the guest list for magicians." "And the wrapping paper?" "Already in the van. Sparkly and shiny, the good stuff." She tapped her clipboard. "Labeled 'Thank You Cards' in case anyone asks." Bill squinted. "Molly, that's a rocket launcher disguised as a birthday cake." "Relax. It's a cake with a surprise inside." "That's not reassuring." A small child wandered over, sticky-faced and holding a balloon. "Are you the robbers?" Molly froze. Bill froze. "Of course not, sweetheart. We're here for the cake." The child considered this, then shrugged and ran off. Molly exhaled. "Okay, go. Grab the big boxes first." "Which ones?" "The ones shaped like ponies. Those are always the expensive ones."
created_at
"2026-10-08T08:26:56.158Z"
10/8/2026, 8:26:56 AM
model
"anthropic/claude-haiku-5.5-20261007"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
727
moderation_latency
(null)
generation_time
3221
tokens_prompt
75
tokens_completion
257
native_tokens_prompt
95
native_tokens_completion
397
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"end_turn"
service_tier
"default"
usage
0.000208
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791448016-2e3apG8WG52amxoP8yqM"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791448016-0Atqg9iuRKs46oDs35wp"
upstream_id
"msg_011CfpTCfGFs6SgjzBVPwqBa"
provider_responses
0
endpoint_id
"b6cdf493-3a61-441f-a134-e7944fddb980"
id
"msg_011CfpTCfGFs6SgjzBVPwqBa"
is_byok
false
latency
727
model_permaslug
"anthropic/claude-haiku-5.5-20261007"
provider_name
"Claude Platform on AWS"
status
200
total_cost
0.000208
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Claude Platform on AWS"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
0.00%
Matches word count
n/a
neededClean
false
words
161
0.00%
Dialogue to Total Word Ratio
Ratio: 68.52%, Deviation: 18.52%
neededClean
false
wordsTotal
162
wordsDialogue
111
0.0004%