NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1GSEZXA740JX1XGDDTRP7WM
Run Details
Model:
Z.AI GLM 5.3 (Reasoning, Max)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
1m 47s
Cost:
$0.015872
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
97.63%
Matches word count
(details)
99.34%
Dialogue to Total Word Ratio
(details)
98.4817%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the hedge. "Bill, why are you wearing a balaclava?" "It's a heist, Molly. Heists require balaclavas." "It's a six-year-old's party." "Professionalism doesn't have an age minimum." "We're stealing presents from children." "Repurposing. We're repurposing presents." "From children, Bill." "You keep saying that like it changes the plan." "Because it should! There will be a clown." "Clowns are just security in makeup." "There's a bouncy castle." "Which we'll avoid. Or use. Depends on the exit strategy." "There is no exit strategy. It's a garden." "Gardens have gates. Gates are exits." "I still say we case the interior first." "We've cased it. It's a living room with a cat." "And the cat?" "Hostile. Possibly armed." "Also, you're allergic to balloons." "One balloon, Molly. One balloon. In 2011." "You swelled up like a parade float." "Which makes me the distraction, doesn't it? I float in, they panic, you bag the loot." "These are children. They're not 'loot.'" "Fine. You bag the... gifts of joy." "Bill, I'm wearing a tutu." "Commitment, Molly. Commitment." Bill adjusted his balaclava. "Why are we even doing this?" "Debts, Molly. Terrible, Sudoku-related debts." "You owe money over Sudoku?" "Don't ask questions I can't answer in a garden." Somewhere, a clown honked. Bill flinched.
created_at
"2026-09-02T10:08:33.715Z"
9/2/2026, 10:08:33 AM
model
"z-ai/glm-5.3-20260816"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1152
moderation_latency
(null)
generation_time
107037
tokens_prompt
75
tokens_completion
3083
native_tokens_prompt
73
native_tokens_completion
3584
native_tokens_completion_images
(null)
native_tokens_reasoning
3237
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0158718
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788343713-rMiFItZHSI0OFTPW4xKJ"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788343713-VItCH1acjtaJqL5oxZSU"
upstream_id
"202609021808336d710d38a3c7468a"
provider_responses
0
endpoint_id
"db975786-e7f6-4838-9e56-3e0b77f1bdc2"
id
"202609021808336d710d38a3c7468a"
is_byok
false
latency
1151
model_permaslug
"z-ai/glm-5.3-20260816"
provider_name
"Z.AI"
status
200
total_cost
0.0158718
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
97.63%
Matches word count
n/a
neededClean
false
words
207
99.34%
Dialogue to Total Word Ratio
Ratio: 92.86%, Deviation: 2.86%
neededClean
false
wordsTotal
210
wordsDialogue
195
98.4817%