NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36NVACHCMAABZC8BCB3XDVE
Run Details
Model:
GPT-6 Luna (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
20.3s
Cost:
$0.001051
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the manor clock struck thirteen, and every guest in the drawing room turned toward the locked conservatory door. Behind its frosted glass, a shadow leaned briefly against the panes, then vanished before Inspector Vale could reach the handle. When he entered, the room smelled of wet roses, extinguished candles, and a metallic odor beneath the heavy velvet curtains. Lord Ashcombe lay beside the piano, one hand around a silver key, his white dinner jacket darkened at the heart. No one screamed; the guests simply watched Vale, as though the dead man had asked them something in the darkness. Vale noticed Lady Wren's gloves were dry, though rain had followed every other guest through the front doors since dusk. Then the clock struck thirteen, and Ashcombe's fingers tightened around the key in Vale's palm, which had begun to tremble. From the gallery above, someone whispered, “You have the wrong key,” though no guest had moved since Vale came inside.
created_at
"2026-09-23T08:24:22.678Z"
9/23/2026, 8:24:22 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
19372
moderation_latency
(null)
generation_time
20262
tokens_prompt
46
tokens_completion
250
native_tokens_prompt
41
native_tokens_completion
2094
native_tokens_completion_images
0
native_tokens_reasoning
1891
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0010511
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790151862-Qt8wBRUWahGXJCQoKrjH"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790151862-W0qpgrnjDipQ4PoSxvqo"
upstream_id
"resp_0db52e416fbc9d99016ab38cb6c47c87d1b15ad80e8809b033"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_0db52e416fbc9d99016ab38cb6c47c87d1b15ad80e8809b033"
is_byok
false
latency
1455
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0010511
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20