NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HNY2MRRCJ4RP6AGBRV6D1
Run Details
Model:
GPT-6 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
19.1s
Cost:
$0.011872
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the church bells rang thirteen times, and Detective Mara Voss found the mayor dead beneath their silent tower. His hands were folded over a sealed envelope, while fresh snow around his body bore no footprints except her own. Mara had arrived by carriage, yet the tracks behind her vanished abruptly at the gate, as though erased by someone. Inside the envelope, a single photograph showed Mara standing beside the mayor, holding a knife she had never seen before. A clock struck one from the empty bell tower, although its mechanism had been dismantled after last winter's devastating fire. Then a voice rose from behind the locked church doors, asking Mara why she had returned to finish the job. She recognized the voice immediately, though its owner had been buried three days earlier, with Mara watching from the graveside. Above her, the bells began again, and something inside the mayor's coat answered each toll with a muffled wet knock.
created_at
"2026-09-23T16:30:46.364Z"
9/23/2026, 4:30:46 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
5929
moderation_latency
(null)
generation_time
19061
tokens_prompt
46
tokens_completion
662
native_tokens_prompt
41
native_tokens_completion
1179
native_tokens_completion_images
0
native_tokens_reasoning
985
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.011872
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181046-PAQ372sxk23U4yUIdhq4"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181046-6OWARH5lHpZWMcZIDyg8"
upstream_id
"resp_08a612e2925561b0016ab3feb6787887d1a9a5f1f9361ba51f"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_08a612e2925561b0016ab3feb6787887d1a9a5f1f9361ba51f"
is_byok
false
latency
1423
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.011872
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20