NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-07-01M4ASYRR6S1ZCJWWTBPRDZ14R
Run Details
Model:
Mistral Medium 3.5
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
2.9s
Cost:
$0.002191
Provider:
Mistral
Evaluation summary
Score
Evaluator
Details Link
0.21%
Matches word count
(details)
36.79%
Dialogue to Total Word Ratio
(details)
18.5010%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
"Alright, Bill, we’re hitting the Johnson’s birthday party at three. Kids’ll be sugar-crazed, parents distracted—perfect cover." "Genius, Molly. But what’s the play? Grab and run?" "Too messy. We go in as clowns. Balloon animals, face paint—blend right in." "Clowns? I look like a deflated balloon in spandex." "Exactly. No one suspects a sad clown. You distract the moms, I’ll hit the gift table." "And the dad with the camera?" "Tell him you’re a ‘professional party entertainer.’ Then drop a pie on his head." "A pie? Where do we get a pie?" "Steal one from the dessert table. It’s a party, Bill. Think on your feet." "Fine. But if I have to make a balloon giraffe, I’m out." "Just nod and inflate. And for God’s sake, don’t call the kid a ‘little brat’ like last time." "Last time was a misunderstanding. That kid *was* a brat." "Focus. In and out. No brats, no arrests." "Got it. But if I see one more glitter cannon, I’m taking a hostage." "Bill." "Kidding. Mostly."
created_at
"2026-10-07T09:08:49.556Z"
10/7/2026, 9:08:49 AM
model
"mistralai/mistral-medium-3.5-20260430"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
422
moderation_latency
(null)
generation_time
2828
tokens_prompt
75
tokens_completion
267
native_tokens_prompt
81
native_tokens_completion
276
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
"default"
usage
0.0021915
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791364129-rfAc1xIEzfXQObFLtALF"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791364129-DHnJcUOm0Cv4WD76GydQ"
upstream_id
"dc29287547bb448989a61f6f784afcfd"
provider_responses
0
endpoint_id
"d21d316f-d540-4718-b36e-3beff85d0cf9"
id
"dc29287547bb448989a61f6f784afcfd"
is_byok
false
latency
422
model_permaslug
"mistralai/mistral-medium-3.5-20260430"
provider_name
"Mistral"
status
200
total_cost
0.0021915
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Mistral"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
0.21%
Matches word count
n/a
neededClean
false
words
172
36.79%
Dialogue to Total Word Ratio
Ratio: 100.00%, Deviation: 10.00%
neededClean
false
wordsTotal
173
wordsDialogue
173
18.5010%