NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-23-01M36NTB93DJD4NXBCXASZHXA1
Run Details
Model:
GPT-6 Luna
Test:
Language Writing
Scenario:
Character dialogue (German) in a story
Duration:
2.2s
Cost:
$0.000034
Provider:
OpenAI
Evaluation summary
Score
Evaluator
Details Link
100.00%
Parse dialogue
(details)
Model input
System
You are an expert novel author in the cozy mystery genre. Take the following details about characters/locations/... into account: <character name="Melanie"> Melanie is a 30-year old college student on a study abroad program. Born in Germany, she is fluent in German and does not yet know any English. She enjoys landscape photography and is always seen with her camera. </character> <character name="Hunter"> Hunter is a 35-year old detective who is investigating a series of mysterious disappearances in the small town of Alborn, MN. He is a former marine and has a no-nonsense attitude. Hunter has no knowledge of German, not even basic phrases. </character> <location name="Alborn, MN"> Alborne is a small town in the Midwest with a population of 3,000. It is known for its annual apple festival. </location>
User/Human
Write the following scene: Melanie is sitting at a cafe, looking at her camera. Hunter approaches her and asks her about a recent event in the town. She responds troubled and shaken in German, and Hunter looks confused. They talk for a while, until Hunter gets a call and has to leave. Return the dialogue between Melanie and Hunter as a voice sheet: Melanie: "..." Hunter: "..."
Model response
Messages
Meta
AI
Melanie: "..." Hunter: "..."
created_at
"2026-09-23T08:23:50.825Z"
9/23/2026, 8:23:50 AM
model
"openai/gpt-6-luna-20260922"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2146
moderation_latency
(null)
generation_time
2146
tokens_prompt
306
tokens_completion
8
native_tokens_prompt
271
native_tokens_completion
13
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"default"
usage
0.0000336
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.20.0; linux; x64))"
http_referer
(null)
request_id
"req-1790151830-uqGYzEVV3uebTjFXyKt2"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1790151830-KRvfYRInBj6wNal97Kof"
upstream_id
"resp_0780b45bbd68f861016ab38c96f16887d1a887d0ec9696a16c"
provider_responses
0
endpoint_id
"05e94e02-b9c2-4bbb-ba55-4082ee9ad687"
id
"resp_0780b45bbd68f861016ab38c96f16887d1a887d0ec9696a16c"
is_byok
false
latency
1508
model_permaslug
"openai/gpt-6-luna-20260922"
provider_name
"OpenAI"
status
200
total_cost
0.0000336
cache_discount
(null)
upstream_inference_cost
0
provider_name
"OpenAI"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Parse dialogue
n/a
dialogue
0
name
"Melanie"
dialogue
"..."
detectedLang
""
heavyLang
""
scores
reliable
false
passes
true
1
name
"Hunter"
dialogue
"..."
detectedLang
""
heavyLang
""
scores
reliable
false
passes
true