NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36NYR32G7P7HN31GTG25B9P
Run Details
Model:
GPT-6 Luna (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write unattributed dialogue
Duration:
9.2s
Cost:
$0.000183
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Count dialogue tags
(details)
Model input
User/Human
You are an expert novel writer in the cozy mystery genre. Write a scene of about 200 words, with dialogue between two characters. The dialogue should be unattributed, meaning the speaker is not explicitly mentioned. The scene should be set in a small town bakery. Only return the story, nothing else.
Model response
Messages
Meta
AI
The bell above the door gave a tired jingle as rain swept into Marigold’s Bakery. A streak of mud crossed the checkerboard floor and stopped beside the cooling racks. “You said the back door was locked.” “It was. The key was still in the flour tin.” “Then how did someone get in?” A tray of cinnamon rolls steamed between them. Beyond it, the empty display case reflected the gray morning—and a small smear of red icing on its glass. “That’s not icing.” “Don’t touch it.” “I wasn’t going to. I have excellent instincts.” “You once mistook a mouse for a walnut.” “It was dark.” Mabel leaned closer to the glass. The red streak ended in a neat little crescent, like a fingerprint. Outside, the church clock struck eight, though it had been five minutes slow for years. “Who knew about the tin?” “Half the town. I told Mrs. Bell it was a terrible hiding place.” “And she told everyone else.” A gust rattled the windows. The bell jingled again, though no one entered. “Did you hear that?” “I heard the bell.” “The door didn’t open.” Behind the counter, the old brass key slid from the flour tin and landed on the floor. Neither of them moved. Then, from the kitchen, the oven timer began to ring.
created_at
"2026-09-23T08:26:15.016Z"
9/23/2026, 8:26:15 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
4827
moderation_latency
(null)
generation_time
9210
tokens_prompt
78
tokens_completion
467
native_tokens_prompt
67
native_tokens_completion
353
native_tokens_completion_images
0
native_tokens_reasoning
73
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0001832
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790151975-gqitzbDE2uXKF2hn4CvA"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790151975-c6YrF5uLESMCaeAzvf44"
upstream_id
"resp_09afdba1ee03f367016ab38d271a0c87d1a8fa13d657d75bad"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_09afdba1ee03f367016ab38d271a0c87d1a8fa13d657d75bad"
is_byok
false
latency
2581
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0001832
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Count dialogue tags
n/a
neededClean
false
dialogueTags
(empty)