NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36R604ZSCSAC1J48XSM4STY
Run Details
Model:
GPT-6 Luna (Reasoning, High)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
10.8s
Cost:
$0.000588
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the lighthouse keeper found the mayor’s coat floating in the harbor, its pockets weighted with smooth, unfamiliar stones. No body surfaced, but a crimson thread clung to the iron ladder, trembling whenever the tide dragged against the pilings. Inspector Vale arrived before dawn, carrying a suitcase, a revolver, and the private conviction that someone had already begun lying. The keeper watched her examine the coat, then whispered, “He was wearing it when he came back from the dead.” Vale looked up sharply; outside, the lighthouse beam crossed the water and illuminated a second figure standing on the breakwater. It raised one gloved hand, then vanished as the beam turned, leaving only a wet footprint on the stone steps. The mayor’s name was carved into its heel, though he had been buried beneath the yew tree three years earlier. From somewhere below the rocks came three slow knocks, and the keeper began counting them with the terror of recognition.
created_at
"2026-09-23T09:05:09.802Z"
9/23/2026, 9:05:09 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
9782
moderation_latency
(null)
generation_time
10789
tokens_prompt
46
tokens_completion
253
native_tokens_prompt
41
native_tokens_completion
1168
native_tokens_completion_images
0
native_tokens_reasoning
962
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0005881
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1790154309-z6zx0Tb8uSbCEbUiktqC"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790154309-VcYWfcVkbNvoUAKx5vBD"
upstream_id
"resp_03836dcd510fb01d016ab39645e5c087d19e6c5b1de2e32ed3"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_03836dcd510fb01d016ab39645e5c087d19e6c5b1de2e32ed3"
is_byok
false
latency
398
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0005881
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20