NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1GT4F20G5B011G8VN3N84T1
Run Details
Model:
Z.AI GLM 5.3 (Reasoning, Max)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
59.4s
Cost:
$0.007705
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
97.63%
Matches word count
(details)
92.48%
Dialogue to Total Word Ratio
(details)
95.0518%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the blueprint across the kitchen table, weighting the corners with stolen paperweights. "Okay. The party starts at two. We go in at two-fifteen, when the cake comes out. Everyone's looking at the cake." Bill squinted at the blueprint. "This is a Chuck E. Cheese." "It's a map of a Chuck E. Cheese." Molly tapped the paper. "The presents get stacked by the ball pit. Classic tactical error." "And the escape route?" "Through the tunnel maze. It empties into the parking lot." Bill rubbed his jaw, considering. He'd spent six years in prison for a bank vault job, and another two planning a museum heist that collapsed when the museum closed. This felt beneath him. On the other hand, rent was due. "What about the kids?" he asked. "What about them?" "Isn't there, like, a moral issue?" Molly stared at him the way a cat stares at a closed door. "It's Derek's kid's party. Derek owes me four hundred dollars." "So this is a debt collection?" "It's a heist, Bill. Heists have casualties." "They're six." "Then they won't remember us." She folded the blueprint with surgical precision. "Wear something cheerful. Nobody suspects a man in a dinosaur shirt." Bill sighed, already reaching for his dinosaur shirt.
created_at
"2026-09-02T10:20:17.354Z"
9/2/2026, 10:20:17 AM
model
"z-ai/glm-5.3-20260816"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
14663
moderation_latency
(null)
generation_time
59370
tokens_prompt
75
tokens_completion
1659
native_tokens_prompt
73
native_tokens_completion
1728
native_tokens_completion_images
(null)
native_tokens_reasoning
1443
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0077054
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788344417-hdWkWDlZJbSJihc0TVl3"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788344417-fv043X08gogMxRjOeVs6"
upstream_id
"20260902182017f4654ee450274456"
provider_responses
0
endpoint_id
"db975786-e7f6-4838-9e56-3e0b77f1bdc2"
id
"20260902182017f4654ee450274456"
is_byok
false
latency
14663
model_permaslug
"z-ai/glm-5.3-20260816"
provider_name
"Z.AI"
status
200
total_cost
0.0077054
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
97.63%
Matches word count
n/a
neededClean
false
words
207
92.48%
Dialogue to Total Word Ratio
Ratio: 55.29%, Deviation: 5.29%
neededClean
false
wordsTotal
208
wordsDialogue
115
95.0518%