NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-05-01M46BH3QRFTD03CCVP3V9Y9TM
Run Details
Model:
GPT-6.1 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
15.1s
Cost:
$0.005932
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Rain battered Blackthorn Manor all evening, but the scream arrived precisely when every clock inside stopped at eleven seventeen simultaneously. Detective Mara Vale found the hostess sprawled beneath her own portrait, wearing pearls nobody remembered seeing before that dreadful night. Her husband stood beside the fireplace, holding an empty brandy glass, while something dark dripped steadily from his immaculate cuff. The servants insisted every door had remained locked, although muddy footprints crossed the carpet and ended halfway up the staircase. Mara knelt beside the body and noticed a folded invitation tucked between cold fingers, addressed in her own distinctive handwriting. She had never met Lady Blackthorn, never accepted this invitation, and certainly never written the words waiting beneath its seal. Inside the envelope, a single line promised that before sunrise Mara would identify the murderer or become the second victim.
created_at
"2026-10-05T15:39:44.254Z"
10/5/2026, 3:39:44 PM
model
"openai/gpt-6.1-sol-20260929"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10203
moderation_latency
(null)
generation_time
15079
tokens_prompt
46
tokens_completion
501
native_tokens_prompt
41
native_tokens_completion
585
native_tokens_completion_images
0
native_tokens_reasoning
415
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.005932
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791214784-H5ViVanAxPUTrZQbc6Bu"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791214784-8smKCTMWBIXMUNH46v3K"
upstream_id
"resp_0491a05af57d820d016ac3c4c05e3887d195ceeea406a972d7"
provider_responses
0
endpoint_id
"dfb9d5c3-63f9-4263-8b78-4fa97dcb3d5e"
id
"resp_0491a05af57d820d016ac3c4c05e3887d195ceeea406a972d7"
is_byok
false
latency
340
model_permaslug
"openai/gpt-6.1-sol-20260929"
provider_name
"OpenAI"
status
200
total_cost
0.005932
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
7
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20