NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36R050C9MH19395NJ2F6XQB
Run Details
Model:
GPT-6 Luna (Reasoning, High)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
25.7s
Cost:
$0.001563
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly unfolded a party map across the laundromat table, pinning one corner beneath a sock. Molly said, “Party starts at noon; parents always leave the cake unattended.” Bill studied the drawing upside down, nodding as if geography were optional. Bill said, “And we stroll in wearing paper crowns, like responsible uncles.” Outside, a delivery van honked; both thieves ducked behind a potted fern. Molly said, “We take every present, but leave the children their ribbons.” Nobody had seen them, except a tortoise wearing a birthday hat. Bill said, “That sounds generous, unless the ribbons are secretly very valuable.” Molly tapped the route, which crossed the garden, climbed the fence, and entered through the kitchen. Molly said, “The getaway is your bicycle, with its squeaky little bell.” Bill packed gloves, a flashlight, and an enormous empty pillowcase. Bill said, “Nobody suspects quiet cyclists carrying suspiciously large sacks after lunch.” The operation seemed simple. Molly said, “We should bring disguises, preferably matching all the birthday decorations.” Bill said, “I have a clown nose that makes me quite trustworthy.” Molly said, “Then let's rehearse, before the invitation becomes evidence against us.” Bill said, “Agreed, though I may accidentally give the gifts back again.”
created_at
"2026-09-23T09:01:58.323Z"
9/23/2026, 9:01:58 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3997
moderation_latency
(null)
generation_time
25389
tokens_prompt
75
tokens_completion
589
native_tokens_prompt
66
native_tokens_completion
3112
native_tokens_completion_images
0
native_tokens_reasoning
2834
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0015626
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1790154118-xKtgCmceEkgQ4DK4wV79"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790154118-Gwg53b99Vg9p9U6TllWB"
upstream_id
"resp_0b99fa625c556b97016ab395868efc87d1976076567afd929a"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_0b99fa625c556b97016ab395868efc87d1976076567afd929a"
is_byok
false
latency
1164
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0015626
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 50.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
100
100.0000%