NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HQA3XWS8WDV0DQEHXKSGR
Run Details
Model:
GPT-6 Sol
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
5.5s
Cost:
$0.001332
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
92.13%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, the lighthouse keeper found a stranger dead beneath the lens, with seawater pooling around his polished leather shoes. The keeper touched the cold brass railing, then noticed fresh blood on its underside and a key inside the stranger's fist. Outside, the foghorn sounded once, though its mechanism had been dismantled that morning, and every vessel lay silent in harbor. He climbed toward the lamp, counting stairs until he heard someone above him whisper the dead man's name: Inspector Vale. The lamp ignited without warning, and across the glass appeared a message written backward in salt: Do not trust her.
created_at
"2026-09-23T16:31:31.466Z"
9/23/2026, 4:31:31 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3807
moderation_latency
(null)
generation_time
5496
tokens_prompt
46
tokens_completion
157
native_tokens_prompt
41
native_tokens_completion
125
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.001332
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181091-1gh2C3A3Ou0tQWZc1gME"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181091-bZOGWih0gsIjzJmt3fqS"
upstream_id
"resp_03b4c96d02a4f98e016ab3fee39eb887d1b9087cbc645a18aa"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_03b4c96d02a4f98e016ab3fee39eb887d1b9087cbc645a18aa"
is_byok
false
latency
2312
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.001332
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
92.13%
Matches word count
n/a
neededClean
false
sentences
5
wordCounts
0
20
1
21
2
20
3
20
4
20