NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-05-01M46BNH12CC4T6WMBKJ28WS7V
Run Details
Model:
GPT-6.1 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
27.7s
Cost:
$0.010482
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Rain battered the windows of Blackthorn Hall as Inspector Vale arrived, carrying a telegram that smelled faintly of burned lavender. Inside, seven guests stood around the dining table, staring at their host, whose untouched soup had already begun to cool. Sir Edmund sat upright beneath the chandelier, a pearl hatpin protruding from his throat like an accusation nobody would acknowledge. The housekeeper insisted every door had remained locked since dinner, although muddy footprints crossed the carpet toward an empty fireplace. Vale removed his gloves slowly, noticing that each guest watched his hands rather than the dead man beside the candles. Above them, the grandfather clock struck thirteen, and Lady Blackthorn whispered that it had stopped working three long winters ago. He unfolded the telegram again, reading the message Sir Edmund had sent six hours before his death was first discovered. The message named no suspect, offered no explanation, and ended with a warning written in a different, distinctly childish hand. “Whatever you are told tonight, Inspector, do not believe my wife when she says that she has never met you.” Lady Blackthorn lifted her veil, and Vale recognized the woman whose funeral he had attended on a cloudless April morning.
created_at
"2026-10-05T15:42:08.958Z"
10/5/2026, 3:42:08 PM
model
"openai/gpt-6.1-sol-20260929"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
13348
moderation_latency
(null)
generation_time
27592
tokens_prompt
46
tokens_completion
634
native_tokens_prompt
41
native_tokens_completion
1040
native_tokens_completion_images
0
native_tokens_reasoning
799
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.010482
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791214928-bYgQqqEEz1IA82LZmhoG"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791214928-OVv5mMYvARRNs7rw61zU"
upstream_id
"resp_0d7ef2dcd4d56370016ac3c55119a887d192c4dfba322ac4c6"
provider_responses
0
endpoint_id
"dfb9d5c3-63f9-4263-8b78-4fa97dcb3d5e"
id
"resp_0d7ef2dcd4d56370016ac3c55119a887d192c4dfba322ac4c6"
is_byok
false
latency
389
model_permaslug
"openai/gpt-6.1-sol-20260929"
provider_name
"OpenAI"
status
200
total_cost
0.010482
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
10
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20
8
20
9
20