NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1GQSF741S12328WH48TWX37
Run Details
Model:
Z.AI GLM 5.3 (Reasoning, Max)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
4m 51s
Cost:
$0.056809
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.9985%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
The stakeout van sat three houses down from the party, windshield plastered with flyers for a clown service nobody remembered hiring. Molly lowered her binoculars—two toilet paper tubes taped together—and pretended they worked. "Status report," she said. Bill consulted his notes, which were a napkin. "Twenty kids. One bouncy castle. Pony on premises. Presents under a dinosaur-shaped table." "A pony." Molly whistled, low and criminal. "We're in the big leagues now." "Remember the ferret store?" Bill shuddered. Some scars never heal, especially ferret-shaped ones. "That ferret had it coming. Focus." She unfolded the map, a crayon drawing bribed off the birthday boy for one juice box. "We go in as caterers during cake. Everyone sings, we load presents into the wagon." "A wagon. Into a party swarming with four-foot detectives." "Kids aren't detectives." "Kids are nothing but detectives. One notices a new shoelace and boom—interrogation." Bill slumped against the horn, which honked mournfully. Somewhere down the street, a dog answered back. He took this as a sign. "We're stealing from children," he muttered. "Children whose parents bought a pony. We're basically Robin Hood." "Robin Hood never fled on a Big Wheel." "He would've, if he'd tried one."
created_at
"2026-09-02T09:39:19.918Z"
9/2/2026, 9:39:19 AM
model
"z-ai/glm-5.3-20260816"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1030
moderation_latency
(null)
generation_time
290910
tokens_prompt
75
tokens_completion
10350
native_tokens_prompt
73
native_tokens_completion
12888
native_tokens_completion_images
(null)
native_tokens_reasoning
12598
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.0568094
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788341959-B7Hzt2JhKOe2rxogHL2x"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788341959-0XDS1W3bK5orZKfZ5C6A"
upstream_id
"20260902173920bc74c5ad7e514735"
provider_responses
0
endpoint_id
"db975786-e7f6-4838-9e56-3e0b77f1bdc2"
id
"20260902173920bc74c5ad7e514735"
is_byok
false
latency
1030
model_permaslug
"z-ai/glm-5.3-20260816"
provider_name
"Z.AI"
status
200
total_cost
0.0568094
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 50.74%, Deviation: 0.74%
neededClean
false
wordsTotal
203
wordsDialogue
103
99.9985%