NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1GRW4RWVTMEHXF7VTWPX1P1
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Max)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
58.7s
Cost:
$0.000617
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.9995%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
"Walk me through it one more time," Molly said. "The target is a birthday party. Saturday. Tommy Henderson, turning six." "Six." "He's loaded. Grandma's flying in from Boca." "Security?" "Bouncy castle, one clown, balloon arch." "Balloon arch?" "Training balloons. Mean ones." "Location?" "Backyard. Ten kids, two parents, one golden retriever." "The dog." "Beagle." "Worse." "Exit strategy?" "Pinata. I rig the candy. We swap the presents for a napping toddler." "There's a napping toddler?" "There's always a napping toddler." "Bill, why are we doing this?" "Retirement fund. Also, Tommy got a pony last year." "A pony?" "Nobody gets a pony at six and stays humble. He's got it coming." "The take?" "Twelve presents, one dinosaur cake, and if we're lucky, Grandma's jewelry." "And the clown?" "The clown sees everything." "I handle clowns. I have a history." "Good history?" "Ongoing." "Fine. Cake and go, nine minutes in and out." "Eight if the dog's asleep." "One more thing. We wear mascot suits. I got us a Buzz Lightyear and a raccoon." "Why a raccoon?" "He was on sale." Molly stared at the blueprints taped to the kitchen wall, a crayon drawing of the yard. "Is the pony thing even true?" Bill paused. "Define true."
created_at
"2026-09-02T09:58:16.102Z"
9/2/2026, 9:58:16 AM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1312
moderation_latency
(null)
generation_time
58649
tokens_prompt
75
tokens_completion
2121
native_tokens_prompt
73
native_tokens_completion
2446
native_tokens_completion_images
(null)
native_tokens_reasoning
2132
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000616975
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788343096-ETcgaNhPlqkW8ExeYT9z"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788343096-zmZd3k8GkX3jRiTc6usO"
upstream_id
"2026090217581664247ddcb17f4d43"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"2026090217581664247ddcb17f4d43"
is_byok
false
latency
1312
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000616975
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
201
100.00%
Dialogue to Total Word Ratio
Ratio: 90.05%, Deviation: 0.05%
neededClean
false
wordsTotal
201
wordsDialogue
181
99.9995%