NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1H5B8M87RMCC91AB5CB2TN5
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
10.7s
Cost:
$0.000076
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
98.71%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
49.3562%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the bounce house, adjusting her cat mask. "Okay, remind me why we're robbing a six-year-old's birthday party." "Because the kid's father is a diamond smuggler, and the loot's hidden in the presents," Bill whispered, wearing a clown nose that kept sliding off. "It's the perfect cover." "There's a pony here, Bill. I did not factor in a pony." "The pony is irrelevant. Stick to the plan. I distract the magician, you swap the gift boxes with the decoys." "The decoys are empty. They'll notice." "Not if the kid's drowning in presents already. Nobody counts at a birthday party." "The mother is counting, Bill. Mothers always count. It's their superpower." "Fine. Then grab the ones with the blue ribbon and leave the rest." "What if the blue ribbon is just... a ribbon?" Bill paused. "It's not." "How do you know?" "Because I spent three weeks bribing a party planner, and I'd like to see that effort pay off." Molly sighed as a stampede of children screamed past them toward the cake. "If a pony steps on me, I'm blaming you." "Just smile and wave like a normal clown." "You're the clown, Bill. I'm a cat. Cats don't wave." "Then be a sneaky cat. Go."
created_at
"2026-09-02T13:36:14.481Z"
9/2/2026, 1:36:14 PM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
4109
moderation_latency
(null)
generation_time
10692
tokens_prompt
75
tokens_completion
313
native_tokens_prompt
73
native_tokens_completion
296
native_tokens_completion_images
(null)
native_tokens_reasoning
8
native_tokens_cached
64
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000075635
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788356174-XoPXHIv6eQ1gaP0v5elV"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788356174-4Ksi5Wwd2vhk0NJMNfTh"
upstream_id
"20260902213614e053444824694092"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"20260902213614e053444824694092"
is_byok
false
latency
4109
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000075635
cache_discount
0.00000768
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
98.71%
Matches word count
n/a
neededClean
false
words
206
0.00%
Dialogue to Total Word Ratio
Ratio: 83.17%, Deviation: 33.17%
neededClean
false
wordsTotal
208
wordsDialogue
173
49.3562%