NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36R73DG44N77J4G61FFKHSX
Run Details
Model:
GPT-6 Luna (Reasoning, High)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
13.6s
Cost:
$0.000619
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the clock above Bellweather Station stopped, though every witness swore its hands had continued moving through the storm. Detective Mara Venn arrived carrying no umbrella, her coat dry, her expression colder than the rain outside the locked doors. Inside, the stationmaster lay beneath the timetable, one gloved hand clutching a ticket dated tomorrow morning, marked in red ink. Nobody admitted seeing him enter, yet fresh footprints crossed the platform and ended abruptly beside Mara’s shoes without continuing farther. Mara studied them, recognizing the narrow heel marks from a pair of boots she had buried beneath her childhood home. Then the station clock struck thirteen, and somewhere beneath the platform, a dead man answered with three measured knocks softly.
created_at
"2026-09-23T09:05:45.913Z"
9/23/2026, 9:05:45 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
9445
moderation_latency
(null)
generation_time
13558
tokens_prompt
46
tokens_completion
490
native_tokens_prompt
41
native_tokens_completion
1229
native_tokens_completion_images
0
native_tokens_reasoning
1072
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0006186
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1790154345-mOoJUt0A6c5YbVvRELDA"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790154345-sW7piamMjDrNNCq338vE"
upstream_id
"resp_071c6e33e9d04ff3016ab3966a08fc87d1bbaffe76f059301f"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_071c6e33e9d04ff3016ab3966a08fc87d1bbaffe76f059301f"
is_byok
false
latency
804
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0006186
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
6
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20