NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-10-07-01M4ATZ4JCTHFXWBD68QTNJ05Q
Run Details
Model:
Mistral Large 4.0 (Reasoning)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
2m 50s
Cost:
$0.029456
Provider:
Mistral
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the crayon-colored invitations across the kitchen table, her black fingerless gloves tapping nervously against the plastic laminate. "Party starts at two," she said. Bill adjusted his tight black tactical ski mask, already sweating profusely in the humid August heat. "We bring bags?" he asked, wiping his brow. They had scoped the suburban location thoroughly: a rented indoor bounce house facility in the suburbs, thirty screaming kids, and a large mountain of brightly wrapped boxes strategically placed right near the emergency exit. The security consisted of one exhausted mother checking her smartphone while holding a lukewarm coffee cup. Molly pulled out a digital timer and set it precisely for twelve long minutes. "Thirty seconds to grab everything," she whispered. Bill nodded solemnly, though his stomach growled audibly in the silence. He pointed toward the window with a trembling finger, spotting the entertainment staff arriving early. "The clown looks terrifying," he mouthed silently. "But very easily distracted," Molly replied, checking her digital watch. Their plan remained flawless: enter during the Happy Birthday song, use the piercing noise to mask the crinkling of wrapping paper, escape through the rear exit disguised as inflatable balloon animals. It was almost too easy. Almost.
created_at
"2026-10-07T09:26:30.232Z"
10/7/2026, 9:26:30 AM
model
"mistralai/mistral-large-4-0-20261006"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1994
moderation_latency
(null)
generation_time
169780
tokens_prompt
75
tokens_completion
11815
native_tokens_prompt
67
native_tokens_completion
14072
native_tokens_completion_images
(null)
native_tokens_reasoning
11484
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
"default"
usage
0.02945604
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.21.0; linux; x64))"
http_referer
(null)
request_id
"req-1791365190-F5k6Gb4rrj80qKAmNeZk"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1791365190-r4QfqLhFFGW7BAQaXz01"
upstream_id
"86b67c4b91ac463cb10a71d11b51fd6c"
provider_responses
0
endpoint_id
"19b3b5d7-23f6-4b63-8109-64a6cc37d996"
id
"86b67c4b91ac463cb10a71d11b51fd6c"
is_byok
false
latency
1983
model_permaslug
"mistralai/mistral-large-4-0-20261006"
provider_name
"Mistral"
status
200
total_cost
0.02945604
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Mistral"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 9.95%, Deviation: 0.05%
neededClean
false
wordsTotal
201
wordsDialogue
20
100.0000%