NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-05-01M46B80XDYNEG6NRAFR5TTBC0
Run Details
Model:
GPT-6.1 Sol (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 10 words each
Duration:
20.4s
Cost:
$0.006842
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 10 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Rain hammered Blackthorn Manor while midnight swallowed the surrounding moors. Inspector Vale arrived late, carrying flowers instead of his revolver. The invitation had promised dinner, reconciliation, and one final confession. Instead, Lord Blackthorn lay beneath the chandelier, his throat opened. Six guests watched silently as blood darkened the Persian carpet. Nobody screamed, which troubled Vale more than the corpse itself. Beside the dead man's hand, a silver bell gleamed wetly. Vale recognized it immediately, though recognition should have been impossible. He had buried that bell with his daughter last winter. Across the room, Lady Blackthorn stood clutching a broken watch. “The murderer is upstairs,” she whispered, pointing toward the nursery. From above came three footsteps, followed by a child's laughter.
created_at
"2026-10-05T15:34:46.452Z"
10/5/2026, 3:34:46 PM
model
"openai/gpt-6.1-sol-20260929"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10424
moderation_latency
(null)
generation_time
20390
tokens_prompt
46
tokens_completion
506
native_tokens_prompt
41
native_tokens_completion
676
native_tokens_completion_images
0
native_tokens_reasoning
516
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.006842
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791214486-A3ZQCiqu9JxCKlfsDlvb"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791214486-aPDWpq4ZXVywe2LPbkBF"
upstream_id
"resp_09bd392777d925f0016ac3c396a8a487d1ba0f74adbfdc109e"
provider_responses
0
endpoint_id
"dfb9d5c3-63f9-4263-8b78-4fa97dcb3d5e"
id
"resp_09bd392777d925f0016ac3c396a8a487d1ba0f74adbfdc109e"
is_byok
false
latency
451
model_permaslug
"openai/gpt-6.1-sol-20260929"
provider_name
"OpenAI"
status
200
total_cost
0.006842
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
12
wordCounts
0
10
1
10
2
10
3
10
4
10
5
10
6
10
7
10
8
10
9
10
10
10
11
10