NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36NYJ0NFBM3KATJPA4WSBC3
Run Details
Model:
GPT-6 Luna
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
3.9s
Cost:
$0.000064
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
49.67%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the lighthouse bell rang twice, though its keeper had been dead for seven hours and remained locked inside. Mara Venn arrived carrying a bloodstained umbrella, despite the cloudless sky and the dry gravel path beneath her boots. On the threshold, Inspector Vale noticed wet footprints leading inward, but none returning through the dust toward the harbor. The victim’s watch had stopped at twelve, while every clock in town insisted it was eleven seventeen exactly. Then the telephone rang upstairs, and Mara whispered, “Don’t answer; he always calls before someone disappears from here.”
created_at
"2026-09-23T08:26:08.795Z"
9/23/2026, 8:26:08 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2756
moderation_latency
(null)
generation_time
3846
tokens_prompt
46
tokens_completion
157
native_tokens_prompt
41
native_tokens_completion
120
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0000641
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790151968-iwCORNbONxLHSefJuGiF"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790151968-ST2myHWjTwvcZU4J8vOk"
upstream_id
"resp_0068f83c183d2af8016ab38d20e4b887d195befcadb247612d"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_0068f83c183d2af8016ab38d20e4b887d195befcadb247612d"
is_byok
false
latency
1469
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0000641
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
49.67%
Matches word count
n/a
neededClean
false
sentences
5
wordCounts
0
20
1
19
2
19
3
18
4
18