NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HZJGBT9S1AF85WS56DTWN
Run Details
Model:
GPT-6 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
17.1s
Cost:
$0.012232
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
92.31%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Rain hammered the conservatory roof when Inspector Vale arrived late. Inside, the orchids leaned toward a corpse beneath glass panes. Sir Edmund's throat bore one narrow, astonishingly clean red wound. His wife stood nearby, clutching a dry green garden hose. She claimed the gardener had locked every door yesterday evening. But muddy footprints crossed the floor without approaching him once. Vale noticed one print stopped beside a shattered wall clock. Its hands pointed toward midnight, though dawn filled everything outside. From the rafters came a faint, uneven clicking sound above. The widow flinched before the inspector raised his head slowly. Something metallic swung overhead, catching the first pale morning light. It was a key, tied carefully to black silk thread. Behind him, the gardener whispered, “That key belongs upstairs, only.” Upstairs, behind locked doors, someone began screaming Sir Edmund's name.
created_at
"2026-09-23T16:36:02.196Z"
9/23/2026, 4:36:02 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
6508
moderation_latency
(null)
generation_time
17075
tokens_prompt
46
tokens_completion
509
native_tokens_prompt
41
native_tokens_completion
1215
native_tokens_completion_images
0
native_tokens_reasoning
1034
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.012232
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181362-QQZsKkvK4kVmfBUOWPw3"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181362-E1GUFeM2H71NVPrwmNcI"
upstream_id
"resp_094190186498f3ba016ab3fff24d7487d1b6862f33b6359794"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_094190186498f3ba016ab3fff24d7487d1b6862f33b6359794"
is_byok
false
latency
1613
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.012232
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
92.31%
Matches word count
n/a
neededClean
false
sentences
13
wordCounts
0
10
1
10
2
10
3
10
4
10
5
10
6
10
7
10
8
10
9
10
10
10
11
10
12
20