NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1H5KX3VVBFJ5HY3A1YQF39Z
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Low)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
12.2s
Cost:
$0.000070
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
99.74%
Matches word count
(details)
36.79%
Dialogue to Total Word Ratio
(details)
68.2661%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly studied the bounce house from across the cul-de-sac, binoculars pressed to her face like a birdwatcher with a criminal record. The objective sat in a glittering pile beneath a banner reading HAPPY BIRTHDAY ZOE. Two hundred pounds of wrapped merchandise, minimum, if her estimate of the swimming-pool-sized gift table held. Bill crouched beside her, wearing a disguise he insisted was bulletproof: cargo shorts, a Hawaiian shirt, and a fake mustache that kept sliding when he sweat. The plan was elegant. Molly would enter as a balloon artist, twisting latex into dachshunds until the moment came. Bill would trigger the smoke machine he'd rigged to the piƱata, plunging the yard into sweet-smelling chaos. In the confusion they'd sweep the presents into a wagon painted to look like aFriendly neighborhood recycling bin. What could go wrong was extensive. Molly had counted nine possible disasters, beginning with the cake table and ending with Zoe's grandmother, who reportedly had a black belt and trust issues. Bill was already greasing the wagon wheels. They had exactly eleven minutes before the magician arrived, and with him, an audience of thirty sugar-armed children who would notice anything missing before it finished vanishing.
created_at
"2026-09-02T13:40:57.601Z"
9/2/2026, 1:40:57 PM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1475
moderation_latency
(null)
generation_time
12129
tokens_prompt
75
tokens_completion
331
native_tokens_prompt
73
native_tokens_completion
274
native_tokens_completion_images
(null)
native_tokens_reasoning
15
native_tokens_cached
64
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000070135
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788356457-2boJcDMwhSUDxOXHJvtp"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788356457-0SKWJjptvZlBUY0iV1j9"
upstream_id
"2026090221405700479e8ad09e463e"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"2026090221405700479e8ad09e463e"
is_byok
false
latency
1474
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000070135
cache_discount
0.00000768
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
99.74%
Matches word count
n/a
neededClean
false
words
196
36.79%
Dialogue to Total Word Ratio
Ratio: 0.00%, Deviation: 10.00%
neededClean
false
wordsTotal
202
wordsDialogue
0
68.2661%