NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HPF5QG0CANET4VTAJH32X
Run Details
Model:
GPT-6 Sol
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
4.8s
Cost:
$0.001472
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
80.33%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
At midnight, Detective Mara Vale arrived at the manor, where every clock had stopped at eleven minutes past nine. The victim lay beneath the grand staircase, one hand clutching a playing card, while rain tapped against the windows. Mara counted thirteen guests in the drawing room, though the dinner register listed fourteen names beneath the date's red ink. No one admitted knowing the missing guest, and no one could explain why his coat was hanging beside the door. Then the youngest maid whispered that the card belonged to a deck she had locked in the study yesterday afternoon. Mara turned it over and found, written across the back, the exact time that all the clocks had stopped.
created_at
"2026-09-23T16:31:03.874Z"
9/23/2026, 4:31:03 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3244
moderation_latency
(null)
generation_time
4751
tokens_prompt
46
tokens_completion
174
native_tokens_prompt
41
native_tokens_completion
139
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.001472
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790181063-cApuLeRG1XuzzgWHrohs"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790181063-sP5d8pNnRcvh3aM9DGpX"
upstream_id
"resp_0def37f130d96729016ab3fec8175c87d190b05487abe14000"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_0def37f130d96729016ab3fec8175c87d190b05487abe14000"
is_byok
false
latency
1423
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.001472
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
80.33%
Matches word count
n/a
neededClean
false
sentences
6
wordCounts
0
19
1
19
2
20
3
20
4
20
5
19