NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-06-01M48W25R1ADKBCJYQF1MSV67Q
Run Details
Model:
Mistral Large 4.0
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
16.7s
Cost:
$0.000712
Provider:
Mistral
Evaluation summary
Score
Evaluator
Details Link
43.38%
Matches word count
(details)
38.38%
Dialogue to Total Word Ratio
(details)
40.8783%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly adjusted her party hat and whispered, "Bill, are you absolutely certain about this?" Bill tightened his clown nose. "The Johnson twins are turning eight. Eight-year-olds have terrible security." "They have a bounce house." "We have a grappling hook." "They have a dad who's a former linebacker." "He's allergic to latex. I checked his Facebook." Molly sighed. "What's our cover story?" "We're the entertainment." "Bill, you can't juggle." "I can juggle if I'm motivated. Also, I'm bringing a whoopee cushion. Distraction." "That's not a distraction, that's an assault." "It's a child. They'll laugh." "They'll sue." Bill leaned closer. "Molly, think about the haul. LEGO sets. Gift cards. One of those robot dogs that costs more than my car." "Your car is a 2003 Honda with a dent shaped like Florida." "Exactly. We need this." Molly straightened her tutu. "Fine. But if we get caught, I'm blaming you." "Naturally. I'm the mastermind." "You're wearing clown shoes." "They're tactical, Molly. Steel-toed." "They squeak." "That's psychological warfare." Molly opened the party door. "After you, mastermind." Bill grinned, adjusted his rainbow wig, and squeaked forward into destiny.
created_at
"2026-10-06T15:07:09.448Z"
10/6/2026, 3:07:09 PM
model
"mistralai/mistral-large-4-0-20261006"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
783
moderation_latency
(null)
generation_time
16646
tokens_prompt
75
tokens_completion
314
native_tokens_prompt
79
native_tokens_completion
315
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
"default"
usage
0.00071207
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791299229-nz7XPpxo6q2CgmMR9zOg"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791299229-ORmp6Rboyq1QUFxY6iJb"
upstream_id
"e2f704f13a374c01a3322383dbdcf8be"
provider_responses
0
endpoint_id
"540e63db-cc7c-4556-bb67-d449579bcf1a"
id
"e2f704f13a374c01a3322383dbdcf8be"
is_byok
false
latency
783
model_permaslug
"mistralai/mistral-large-4-0-20261006"
provider_name
"Mistral"
status
200
total_cost
0.00071207
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Mistral"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
43.38%
Matches word count
n/a
neededClean
false
words
183
38.38%
Dialogue to Total Word Ratio
Ratio: 80.11%, Deviation: 9.89%
neededClean
false
wordsTotal
186
wordsDialogue
149
40.8783%