NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1GT74ZVXWJ9RG07TK0S6W21
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Max)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
28.7s
Cost:
$0.000212
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
98.71%
Matches word count
(details)
52.20%
Dialogue to Total Word Ratio
(details)
75.4541%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
"Walk me through it again," Bill said. "Little Timmy's seventh birthday. Saturday, two p.m. Bouncy castle, pony rides, presents galore." "Security?" "There's a bounce house attendant named Doug. Doesn't take his eyes off the inflatable." "Doug's a pro?" "Doug's twelve, Bill." "Juveniles crack under pressure. The mother?" "Susan. Counts every gift twice, logs who gave what." "Thank-you cards. Diabolical." "Exactly. So we strike during the piƱata. She's blindfolded for that. Goes for the pull-string every year." "The clown? There's always a clown." "The clown's me, Bill." "You're the clown?" "Face paint's already in the van. That's my way in." "So I'm the pony guy." "You're the pony guy." "I don't know any pony tricks." "You eat hay and look sad. That's the whole job." "And the presents?" "Garbage bags in the diaper bin. Nobody checks the diaper bin." "Bill, this is genius." "It's Molly. I'm Bill." "Right. Sorry. The wig's making me jumpy." "You've been wearing the wig since Tuesday." "...It's grown on me." "Take it off before the job, you maniac. Where's the van?" "The van's a problem." "The van's a problem how?" "I told the party planner we were the magic show."
created_at
"2026-09-02T10:21:45.346Z"
9/2/2026, 10:21:45 AM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3653
moderation_latency
(null)
generation_time
28686
tokens_prompt
75
tokens_completion
883
native_tokens_prompt
73
native_tokens_completion
825
native_tokens_completion_images
(null)
native_tokens_reasoning
519
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000211725
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788344505-qG4Cw0FrCCE765UbFeTp"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788344505-yQKxeaZkY7sn1S6s6FoV"
upstream_id
"20260902182145dd2bec98690b429f"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"20260902182145dd2bec98690b429f"
is_byok
false
latency
3653
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000211725
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
98.71%
Matches word count
n/a
neededClean
false
words
194
52.20%
Dialogue to Total Word Ratio
Ratio: 98.98%, Deviation: 8.98%
neededClean
false
wordsTotal
196
wordsDialogue
194
75.4541%