NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-06-01M48VW8WF6DYDVNGSSY1VFD2T
Run Details
Model:
Mistral Large 4.0
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
8.0s
Cost:
$0.000664
Provider:
Mistral
Evaluation summary
Score
Evaluator
Details Link
95.99%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
47.9934%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the party invitation across the kitchen table like it was a treasure map. "Three o'clock, Saturday, at the community center. Fifty kids, unlimited juice boxes, and a gift table the size of a Volkswagen." Bill squinted at the paper. "You want us to rob a children's party?" "We're not robbing children, Bill. We're robbing parents. There's a difference. The kids are just... collateral ambiance." "They're six, Molly. Six-year-olds with sticky fingers and no concept of personal property. They'll hand us the presents." "Exactly. Low security. High reward. I've already got the clown nose." Bill rubbed his temples. "You bought a clown nose." "I rented it. Three hours, refundable deposit. Look, we dress as party entertainers, blend in with the other twelve adults pretending to enjoy magic tricks, and when the cake comes out—distraction—we load the gift table into the van." "And the kids?" "Face paint. We'll be unrecognizable. Besides, they'll think it's part of the show." Molly grinned. "Bill, think of the haul. Video games, tablets, that robot dog everyone's talking about." Bill stared at the invitation. "The robot dog barks when you kick it." "See? Already planning the resale."
created_at
"2026-10-06T15:03:56.088Z"
10/6/2026, 3:03:56 PM
model
"mistralai/mistral-large-4-0-20261006"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
632
moderation_latency
(null)
generation_time
7922
tokens_prompt
75
tokens_completion
310
native_tokens_prompt
79
native_tokens_completion
292
native_tokens_completion_images
(null)
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
"default"
usage
0.000664
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791299036-JSwaLxUeUlq4Jo5103LY"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791299036-CpkFef11FPa1W0SO0s7a"
upstream_id
"9cc1f6b4e0b84284808a64f4cd40773b"
provider_responses
0
endpoint_id
"19b3b5d7-23f6-4b63-8109-64a6cc37d996"
id
"9cc1f6b4e0b84284808a64f4cd40773b"
is_byok
false
latency
630
model_permaslug
"mistralai/mistral-large-4-0-20261006"
provider_name
"Mistral"
status
200
total_cost
0.000664
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Mistral"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
95.99%
Matches word count
n/a
neededClean
false
words
192
0.00%
Dialogue to Total Word Ratio
Ratio: 84.02%, Deviation: 34.02%
neededClean
false
wordsTotal
194
wordsDialogue
163
47.9934%