NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1KCKP17K45NX9K7VKT2XT04
Run Details
Model:
DeepSeek V4 Flash 0731 (Reasoning, High)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
45.5s
Cost:
$0.007029
Provider:
Alibaba
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
The clock struck nine as the butler found the body in the study, cold and motionless on the Persian carpet. A single crimson rose lay beside the victim's outstretched hand, its petals still fresh and untouched by the chaos around. Inspector Hale knelt, studying the scene with a calm, calculating grey-eyed intensity, then noticed the open window and one footprint. The footprint was small, possibly a woman's, leading away from the body's crumpled form and stopping at the heavy door. Hale's jaw tightened; he knew this rose and that footprint from a previous case; the pattern was unmistakably the same. Hale glanced down, saw the dark stain, and realized the butler's accusation was correct, everything stopping suddenly in that moment.
created_at
"2026-09-03T10:21:39.261Z"
9/3/2026, 10:21:39 AM
model
"deepseek/deepseek-v4-flash-20260731"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
904
moderation_latency
(null)
generation_time
45366
tokens_prompt
46
tokens_completion
5524
native_tokens_prompt
118
native_tokens_completion
6617
native_tokens_completion_images
(null)
native_tokens_reasoning
6466
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.007029088
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788430899-3o9WNqUc45lODOgpNBAp"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788430899-1duDzlmGWQrbCW5DBOaW"
upstream_id
"chatcmpl-562d6b3d-fc42-946a-9e55-2a37bee45a5d"
provider_responses
0
endpoint_id
"47a93b18-a9c6-485b-9a47-2afb6ca790dc"
is_byok
false
latency
66
model_permaslug
"deepseek/deepseek-v4-flash-20260731"
provider_name
"BaseTen"
status
429
1
endpoint_id
"f098dac5-389b-4ba6-baa1-327d1c2f73a6"
id
"chatcmpl-562d6b3d-fc42-946a-9e55-2a37bee45a5d"
is_byok
false
latency
904
model_permaslug
"deepseek/deepseek-v4-flash-20260731"
provider_name
"Alibaba"
status
200
total_cost
0.007029088
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Alibaba"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
6
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20