NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36P5JQ9EK287HHJAE82VE5Y
Run Details
Model:
GPT-6 Luna
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
6.8s
Cost:
$0.000068
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
74.87%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Detective Mara Venn arrived before dawn, carrying yesterday’s rain. The manor’s front door stood open, inviting nobody inside. A dead clock ticked softly beside the entrance hall mirror. On its glass, someone had written her name in blood. Behind her, the chauffeur whispered, “You weren’t expected tonight.” Mara turned; the gravel drive stretched empty beneath moonlight. Then every window upstairs went dark, one after another. Somewhere above, a floorboard creaked beneath an unseen footstep. She counted the seconds, listening for another sound. It came from the locked room directly behind her.
created_at
"2026-09-23T08:29:58.895Z"
9/23/2026, 8:29:58 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
5556
moderation_latency
(null)
generation_time
6731
tokens_prompt
46
tokens_completion
157
native_tokens_prompt
41
native_tokens_completion
127
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0000676
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790152198-npCiFJlZrZl50M4F7Ymb"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790152198-LS9QC3zllWBQohilRtbL"
upstream_id
"resp_028a24659d88285d016ab38e0703fc87d19ffe97c37c81eaeb"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_028a24659d88285d016ab38e0703fc87d19ffe97c37c81eaeb"
is_byok
false
latency
4289
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0000676
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
74.87%
Matches word count
n/a
neededClean
false
sentences
10
wordCounts
0
9
1
9
2
10
3
10
4
9
5
9
6
9
7
9
8
8
9
9