NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1K5F1QPPCS7DEC277GNDYVZ
Run Details
Model:
Muse Spark 1.2 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
14.0s
Cost:
$0.015104
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly whispered over the blueprints. "Bill, the target is little Timmy's seventh birthday, noon, lots of presents piled by the cake." Bill adjusted his clown nose nervously. "Right, you sure this is worth it? Stealing from kids feels low, even for us." "You kidding? Those parents overcompensate, the presents are untouched, expensive, easy resale." "Okay, so we go as entertainers? I hate clowns." "Perfect cover, you will juggle, I will do magic, then we grab the loot while they sing happy birthday." "And the parents? They will notice two shady clowns stuffing gifts into duffel bags." "Distraction, I release ten doves, you pop balloons, chaos, we vanish through the back gate quietly." "Brilliant, absolutely terrible morally, but brilliant. Do we get cake?" "Bill, focus, we get presents, then we buy our own cake, a whole bakery." He sighed, tugging at his oversized shoes. "Molly, if Timmy cries, I quit heists forever, I mean it." "Relax, Timmy gets adventure, we get profit, everybody wins, especially the toy store later." "What about the bouncy castle? Can we use it for escape?" "No, Bill, we steal presents not childhoods, stick to the plan, okay?" She grinned. "Fine, no castle, promise? Promise. Let's go."
created_at
"2026-09-03T08:16:47.355Z"
9/3/2026, 8:16:47 AM
model
"meta/muse-spark-1.2-20260805"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3519
moderation_latency
(null)
generation_time
13937
tokens_prompt
75
tokens_completion
357
native_tokens_prompt
68
native_tokens_completion
3534
native_tokens_completion_images
0
native_tokens_reasoning
3185
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.0151045
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788423407-U7jaivOyI9jBYZ3fKVM5"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788423407-fusAsqnyTAmdOaPztMKR"
upstream_id
"resp_6a992cefbfbdd9af5b1243d5"
provider_responses
0
endpoint_id
"ffc8e76e-808a-44d2-bd21-c6ac16e721cb"
id
"resp_6a992cefbfbdd9af5b1243d5"
is_byok
false
latency
513
model_permaslug
"meta/muse-spark-1.2-20260805"
provider_name
"Meta"
status
200
total_cost
0.0151045
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 90.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
180
100.0000%