NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1H5HS5170T3AMY8ANP3GSV1
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
13.9s
Cost:
$0.000074
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
99.98%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
49.9920%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the blueprint across the van's hood and traced a finger along the enemy lines. "The balloons are the alarm system." "Balloon arch," Bill corrected, squinting. "Any kid pops one, confetti falls, parents scream. It's genius, really." "We've stolen diamonds from a panic room. We can handle tissue paper." "That tissue paper is backed by a bounce house, Molly. A BOUNCE HOUSE." Bill threw up his hands. "Do you know what happens to a grown man in a bounce house?" "You find your center and commit." "I find my lunch coming back up, that's what." She rolled the map tighter and smacked him with it. "Focus. Presents go into the duffel at exactly two fifteen, during the magic show." "The magician?" "Great Uncle Randy. He's terrible. Everyone will be watching him bomb." Bill nodded slowly. "And the clown?" Molly froze. "What clown?" "Flyer said there'd be a clown." "Nobody said anything about a clown." She stared at the blueprint like it had betrayed her. "Fine. New plan. I take the clown, you take the cake table." "Why do YOU get the clown?" "Because you cry when clowns sneeze, Bill, and I need my partner functional." He grumbled, loading the duffel anyway.
created_at
"2026-09-02T13:39:48.007Z"
9/2/2026, 1:39:48 PM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
5083
moderation_latency
(null)
generation_time
13868
tokens_prompt
75
tokens_completion
313
native_tokens_prompt
73
native_tokens_completion
291
native_tokens_completion_images
(null)
native_tokens_reasoning
8
native_tokens_cached
64
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000074385
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788356388-EqRpassAuitOMfXaU3QO"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788356388-PuWDjqPygZqtAqsnWHRJ"
upstream_id
"20260902213948a9aca293ae354d25"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"20260902213948a9aca293ae354d25"
is_byok
false
latency
5083
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000074385
cache_discount
0.00000768
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
99.98%
Matches word count
n/a
neededClean
false
words
202
0.00%
Dialogue to Total Word Ratio
Ratio: 72.77%, Deviation: 22.77%
neededClean
false
wordsTotal
202
wordsDialogue
147
49.9920%