NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1KCXQD819E8RVQXT4H7ZH3Z
Run Details
Model:
DeepSeek V4 Pro 0813 (Reasoning, High)
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
1m 3s
Cost:
$0.014713
Provider:
Parasail
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
The manor's grandfather clock struck midnight, echoing through empty halls. Detective Arden Cole stepped over the threshold, rain soaking him. A single candle flickered beside the body near the hearth. No one heard the shot over the storm's violent roar. A woman in a crimson dress knelt close, hands shaking. Cole knelt beside her, studying the wound without touching it. "The bullet entered just below the left eye," she whispered. "Then our killer knew exactly where he was aiming tonight." A brass key lay inches from the dead man's hand. It bore the initial V, polished to a dull gleam. Cole picked it up, turning it toward the flickering light. "V for Victoria," he said, "or perhaps for vengeance tonight." A floorboard creaked above them, and every eye looked upward.
created_at
"2026-09-03T10:27:08.341Z"
9/3/2026, 10:27:08 AM
model
"deepseek/deepseek-v4-pro-20260813"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
715
moderation_latency
(null)
generation_time
63200
tokens_prompt
46
tokens_completion
3621
native_tokens_prompt
118
native_tokens_completion
3676
native_tokens_completion_images
(null)
native_tokens_reasoning
3424
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.01471272
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788431228-RYTefwVb4BeA5cxtjrI0"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788431228-iZjcBl6WqrDvSUtR4ocF"
upstream_id
"4119ccaaed8e44bf9a9e1246d19ae919"
provider_responses
0
endpoint_id
"a69687f5-7209-4d1d-b4d6-50a9939aa5ce"
is_byok
false
latency
105
model_permaslug
"deepseek/deepseek-v4-pro-20260813"
provider_name
"DeepInfra"
status
429
1
endpoint_id
"487a0ef9-ddaa-4d3d-b882-7912daa08555"
id
"4119ccaaed8e44bf9a9e1246d19ae919"
is_byok
false
latency
715
model_permaslug
"deepseek/deepseek-v4-pro-20260813"
provider_name
"Parasail"
status
200
total_cost
0.01471272
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Parasail"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
13
wordCounts
0
10
1
10
2
10
3
10
4
10
5
10
6
10
7
10
8
10
9
10
10
10
11
10
12
10