NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1H5GBWXEN0CQYJ3F17MPV30
Run Details
Model:
Z.AI GLM 5.3 Flash (Reasoning, Low)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
7.1s
Cost:
$0.000051
Provider:
Z.AI
Evaluation summary
Score
Evaluator
Details Link
40.32%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
The rain hammered against the tall windows of Blackwood Manor while Inspector Reeves studied the body on the floor. Lord Edmund Ashworth lay slumped beside his massive oak desk, a silver letter opener protruding from his chest. His wife Margaret stood trembling near the fireplace, her pale hands gripping a silk handkerchief with visible desperation and fear. The housekeeper had discovered the body at precisely seven o'clock, according to the shaking statement she gave the police. Reeves walked slowly around the room, noting the spilled brandy glass, the scattered papers, and the open wall safe. Nothing appeared to be missing from the safe itself, which struck the inspector as deeply strange and suspicious. The storm had knocked out the electricity hours earlier, leaving only flickering candlelight to cast dancing shadows everywhere. Six guests remained inside the manor, stranded by flooded roads, and every single one had motive.
created_at
"2026-09-02T13:39:01.668Z"
9/2/2026, 1:39:01 PM
model
"z-ai/glm-5.3-flash-20260826"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1189
moderation_latency
(null)
generation_time
7101
tokens_prompt
46
tokens_completion
248
native_tokens_prompt
47
native_tokens_completion
190
native_tokens_completion_images
(null)
native_tokens_reasoning
9
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.000051025
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788356341-CUKUrPxa1skpHcuDR03u"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788356341-n356Q67sXD1ANOCwMBhi"
upstream_id
"20260902213901ae3c1aa40a1f4333"
provider_responses
0
endpoint_id
"8e9fe48b-2f91-41c3-a8a7-e4a93a8c4ff0"
id
"20260902213901ae3c1aa40a1f4333"
is_byok
false
latency
1189
model_permaslug
"z-ai/glm-5.3-flash-20260826"
provider_name
"Z.AI"
status
200
total_cost
0.000051025
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Z.AI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
40.32%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
19
1
18
2
20
3
19
4
19
5
18
6
18
7
16