NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36P5FWN970H4PE1HPTAJC9K
Run Details
Model:
GPT-6 Luna (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
18.7s
Cost:
$0.001017
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, Detective Mara Venn found the library door ajar, though the storm had sealed every window against the rain. Inside, a grandfather clock had stopped at eleven, and someone had arranged its hands into a trembling accusation of murder. Lord Ashcombe lay beside the desk, one gloved hand clutching a wet match and the other pointing toward empty shelves. No blood stained the carpet, but a red fingerprint marked the portrait of Ashcombe’s ancestor, whose eyes seemed newly wet. Mara recognized the match from the conservatory, where every candle had been deliberately extinguished before the guests arrived for dinner. Then the butler whispered that only six people possessed keys, and five were gathered downstairs awaiting questions from Detective Venn. From beneath the desk, a telephone rang once, although its disconnected wire curled across the floor like a blackened vein. Mara lifted the receiver and heard her own voice whisper, “You’re late,” as the dead man’s lips began moving silently.
created_at
"2026-09-23T08:29:55.995Z"
9/23/2026, 8:29:55 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
7505
moderation_latency
(null)
generation_time
18419
tokens_prompt
46
tokens_completion
369
native_tokens_prompt
41
native_tokens_completion
2025
native_tokens_completion_images
0
native_tokens_reasoning
1821
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0010166
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790152195-LOGsO5TKbolcG3OF0wu7"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790152195-BUsnooJC1dibm5YwPHUj"
upstream_id
"resp_0910edf163e5f4d6016ab38e04139087d19252da61fd5b813c"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_0910edf163e5f4d6016ab38e04139087d19252da61fd5b813c"
is_byok
false
latency
1182
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0010166
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
8
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20