NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-05-01M46BAQBP3ESW15F89HP6G02W
Run Details
Model:
GPT-6.1 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
23.2s
Cost:
$0.009552
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
By midnight, the rain had erased every footprint outside Blackthorn House, except the one leading away from its locked chapel. Inspector Vale arrived carrying a broken umbrella and the letter his dead wife had apparently posted only three days earlier. Inside, seven guests waited beneath portraits whose painted eyes seemed fixed upon the silver tray resting beside an empty armchair. The tray held a cracked teacup, a wedding ring, and a brass key still wet with something darker than rain. Nobody mentioned the body until the grandfather clock struck thirteen, and the youngest guest began laughing without lifting her eyes. Sir Edmund lay upstairs with his throat cut neatly, although every knife in the house remained locked inside the kitchen. Vale asked who had discovered him, and seven people answered together, each naming somebody who was not in the room. But when he unfolded his wife's letter, the first line made their answers seem suddenly less impossible than his arrival. “My darling, if Edmund is dead when you read this, please remember that I warned you about the eighth guest.”
created_at
"2026-10-05T15:36:14.973Z"
10/5/2026, 3:36:14 PM
model
"openai/gpt-6.1-sol-20260929"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10079
moderation_latency
(null)
generation_time
23142
tokens_prompt
46
tokens_completion
693
native_tokens_prompt
41
native_tokens_completion
947
native_tokens_completion_images
0
native_tokens_reasoning
736
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.009552
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791214574-I1i1Dds17iNRsOvrrCYX"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791214574-ADv1rS8ZF4uswh9KO2ng"
upstream_id
"resp_0016d03e47f20212016ac3c3ef140887d199ef3a3948a2ad1a"
provider_responses
0
endpoint_id
"dfb9d5c3-63f9-4263-8b78-4fa97dcb3d5e"
id
"resp_0016d03e47f20212016ac3c3ef140887d199ef3a3948a2ad1a"
is_byok
false
latency
405
model_permaslug
"openai/gpt-6.1-sol-20260929"
provider_name
"OpenAI"
status
200
total_cost
0.009552
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
9
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20
6
20
7
20
8
20