NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1H5D6T84YRV3996PD5WR9AV
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
13.9s
Cost:
$0.000081
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
99.92%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
49.9595%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the blueprints across the van's hood and tapped a crayon drawing with her fingernail. "Here's the target. Little Timmy's seventh birthday. Fortress-level security, apparently." "A children's party? Molly, we've hit casinos. Banks. That one haunted lighthouse." "And none of them had a bouncy castle. This one does. We steal the presents, we're legends." "Legends? The papers will call us the Present Predators." "I already had the Present Predators. I was thinking the Birthday Bandits." "That's worse." "It's memorable, Bill." "Fine. Walk me through the plan." "We enter through the balloon arch at precisely two p.m., when the clown performs his juggling act. Everyone will be watching." "Except the parents watching their kids." "We'll be disguised as parents." "Parents who come to a stranger's birthday party to steal gifts?" "Exactly. Nobody questions a parent at a party. That's the beauty of it." "What's the exit strategy?" "Bouncy castle. We go in the entrance, out the emergency slide, over the fence, into the getaway minivan." "The one with the 'World's Okayest Dad' bumper sticker?" "That's our cover, yes." Bill sighed. "You know what? After the lighthouse job, this actually sounds reasonable." "That's the spirit. Now help me practice my kids' party small talk."
created_at
"2026-09-02T13:37:18.161Z"
9/2/2026, 1:37:18 PM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1413
moderation_latency
(null)
generation_time
13888
tokens_prompt
75
tokens_completion
345
native_tokens_prompt
73
native_tokens_completion
304
native_tokens_completion_images
(null)
native_tokens_reasoning
8
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000081475
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788356238-E8NRsgA38tkbuoPMOGR1"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788356238-WWbHVGF3EGtKCY7t7vxa"
upstream_id
"20260902213718ab0ec987682e4905"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"20260902213718ab0ec987682e4905"
is_byok
false
latency
1413
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000081475
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
99.92%
Matches word count
n/a
neededClean
false
words
203
0.00%
Dialogue to Total Word Ratio
Ratio: 91.18%, Deviation: 41.18%
neededClean
false
wordsTotal
204
wordsDialogue
186
49.9595%