NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-07-01M4ATVGMQCWN3A6JATGEF28FY
Run Details
Model:
Mistral Medium 3.5
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
3.3s
Cost:
$0.002851
Provider:
Mistral
Evaluation summary
Score
Evaluator
Details Link
90.48%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
45.2419%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly adjusted her Santa hat, the jingle bell dangling precariously. The party raged behind the fence, a riot of sugar-fueled chaos. She nudged Bill. "You sure the dog’s asleep?" Bill squinted at the yard. "Either that or he’s plotting world domination." He adjusted his fake beard. "We go in, grab the loot, out before the piñata explodes." Molly sighed. "Last time, you got stuck in a ball pit." "In my defense, those balls were *sticky*." Bill cracked his knuckles. "This time, we’re pros." A clown blew a horn. Molly flinched. "Pros don’t flinch." "Pros also don’t rob kids," Bill muttered. "Kids have *too* many presents," Molly said. "It’s practically charity." Bill nodded. "Fine. But if we get caught, I’m blaming the reindeer antlers." Molly grinned. "Deal." They crouched, eyeing the back door. A child screamed. Bill swallowed. "Or we could just… not." Molly shoved him forward. "Too late. We’re committed." The door creaked open. Inside, a mountain of wrapped boxes gleamed under twinkling lights. Bill whispered, "Jackpot." Molly grabbed a sack. "Let’s move." A tiny voice piped up. "Who are you?" They froze. A toddler stared, clutching a half-eaten cupcake. Bill forced a laugh. "Uh… surprise entertainers?" The toddler hurled the cupcake at his face. Molly groaned. "We’re *so* fired."
created_at
"2026-10-07T09:24:31.52Z"
10/7/2026, 9:24:31 AM
model
"mistralai/mistral-medium-3.5-20260430"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
204
moderation_latency
(null)
generation_time
3201
tokens_prompt
75
tokens_completion
347
native_tokens_prompt
81
native_tokens_completion
364
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
"default"
usage
0.0028515
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791365071-ydOqC8OkobPaRyCjzRup"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791365071-oPjdASnk1mLpSUNrKpsg"
upstream_id
"8fa65caa45394505807782eba58bc590"
provider_responses
0
endpoint_id
"d21d316f-d540-4718-b36e-3beff85d0cf9"
id
"8fa65caa45394505807782eba58bc590"
is_byok
false
latency
204
model_permaslug
"mistralai/mistral-medium-3.5-20260430"
provider_name
"Mistral"
status
200
total_cost
0.0028515
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Mistral"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
90.48%
Matches word count
n/a
neededClean
false
words
210
0.00%
Dialogue to Total Word Ratio
Ratio: 43.40%, Deviation: 33.40%
neededClean
false
wordsTotal
212
wordsDialogue
92
45.2419%