NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M37HMC1V1R4NAW3MAQE7XKVD
Run Details
Model:
GPT-6 Sol
Test:
Dialogue tags
Scenario:
Write unattributed dialogue
Duration:
7.2s
Cost:
$0.002844
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Count dialogue tags
(details)
Model input
User/Human
You are an expert novel writer in the cozy mystery genre. Write a scene of about 200 words, with dialogue between two characters. The dialogue should be unattributed, meaning the speaker is not explicitly mentioned. The scene should be set in a small town bakery. Only return the story, nothing else.
Model response
Messages
Meta
AI
By nine o’clock, the bakery smelled of cinnamon, warm butter, and rain drying off wool coats. Outside, the bell on the town hall struck once. Inside, someone had stolen the silver cake server. “You’re sure it was here?” “Beside the lemon tarts. I polished it this morning.” “Maybe Mrs. Bell borrowed it.” “Mrs. Bell borrows recipes. She returns those, eventually.” A tray of currant buns cooled on the counter. Beneath it lay a dusting of flour, interrupted by one neat oval print. “That isn’t a shoe mark.” “No. It’s the base of the honey jar.” “The jar’s over there.” “And its lid is sticky.” Rain tapped at the window. Across the square, the florist was sweeping water from his doorstep, while a customer in a red scarf waited under the awning. “You think our thief stopped for honey?” “I think our thief put something inside the jar.” The lid came loose with a soft pop. There was no silver cake server, only a folded note, sweet at the corners. “‘Look behind the third sack of flour.’” “That’s where I hid it.” “Then who wrote the note?” From the kitchen came the scrape of a chair. Both looked toward the swinging door as it settled gently into place.
created_at
"2026-09-23T16:29:55.137Z"
9/23/2026, 4:29:55 PM
model
"openai/gpt-6-sol-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2816
moderation_latency
(null)
generation_time
7139
tokens_prompt
78
tokens_completion
328
native_tokens_prompt
67
native_tokens_completion
271
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.002844
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790180995-tSWsX4EiAI6sX9xbT7gT"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790180995-zgFQSambnra1TXXr2q9x"
upstream_id
"resp_0a3ced2fdf7cd312016ab3fe833dcc87d1a2ff1fd83aa4961c"
provider_responses
0
endpoint_id
"1cfc7d9d-4404-4b8e-9ee4-58ae45c9dcd4"
id
"resp_0a3ced2fdf7cd312016ab3fe833dcc87d1a2ff1fd83aa4961c"
is_byok
false
latency
1117
model_permaslug
"openai/gpt-6-sol-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.002844
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Count dialogue tags
n/a
neededClean
false
dialogueTags
(empty)