NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HBWXFFSM03N7Z1SSXVQKS
Run Details
Model:
GPT-6 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
34.3s
Cost:
$0.012092
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the lighthouse blinked twice, then never blinked again. Detective Mara Vale arrived before the retreating tide revealed everything. The harbor master smashed the locked door with his lantern. Inside, the keeper sat dead beside his untouched birthday supper. His watch ticked steadily, though someone had removed its hands. Across the table lay three keys, each stained with salt. Only one opened the lighthouse, according to the harbor master. But the door had been bolted from within since dusk. Mara noticed wet footprints leading upstairs, never coming down again. The final print ended beneath a portrait of her father. Her father had drowned here twenty years earlier, officially alone. Behind the glass, a handwritten message read: “He never drowned.”
created_at
"2026-09-23T16:25:17.582Z"
9/23/2026, 4:25:17 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
8301
moderation_latency
(null)
generation_time
34124
tokens_prompt
46
tokens_completion
885
native_tokens_prompt
41
native_tokens_completion
1201
native_tokens_completion_images
0
native_tokens_reasoning
1034
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.012092
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790180717-FWVHBr9GuJs7kYOlMNqT"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790180717-Se5lzo9tDIREG0xep6mT"
upstream_id
"resp_0e9da154e1d9bb35016ab3fd6ddb6c87d1a5bb2055573739e2"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_0e9da154e1d9bb35016ab3fd6ddb6c87d1a5bb2055573739e2"
is_byok
false
latency
1056
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.012092
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
12
wordCounts
0
10
1
10
2
10
3
10
4
10
5
10
6
10
7
10
8
10
9
10
10
10
11
10