NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36NVBWFTDZAE2YW0AEYFRAX
Run Details
Model:
GPT-6 Luna (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
14.2s
Cost:
$0.000700
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the clock above Bellweather Station stopped, though every other clock in town continued counting toward morning without explanation. Mara Venn found the stationmaster face down beneath the timetable, one gloved hand clutching a ticket dated exactly tomorrow morning. Rain tapped the skylights, and somewhere in the locked building a telephone rang three times before falling silent without answer. Mara checked the doors; each was bolted from inside, and the stationmaster’s wet coat hung neatly beside the entrance door. She turned the ticket over and found her own name printed beneath a destination no train had served in decades. Behind her, the station clock began ticking backward, while the dead man’s pocket watch continued advancing toward midnight steadily, inexorably. Then the stationmaster’s eyes opened, and he whispered, “Don’t let the conductor see you,” before dying for the second time.
created_at
"2026-09-23T08:24:24.215Z"
9/23/2026, 8:24:24 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
13108
moderation_latency
(null)
generation_time
14156
tokens_prompt
46
tokens_completion
239
native_tokens_prompt
41
native_tokens_completion
1392
native_tokens_completion_images
0
native_tokens_reasoning
1209
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0007001
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790151864-72mwdI45OkVcfWwKyHos"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790151864-fv3VMeXSqo3bRO2J9Tht"
upstream_id
"resp_0dc4a6895d59c88c016ab38cb8509887d1a293063074f8ab2f"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_0dc4a6895d59c88c016ab38cb8509887d1a293063074f8ab2f"
is_byok
false
latency
590
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0007001
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
7
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20