NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1K5CTJ42P96F1NV6KVG090J
Run Details
Model:
Muse Spark 1.3 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
1m 24s
Cost:
$0.020863
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the minivan, peering over bright balloons toward the noisy backyard birthday chaos. "Did you bring the clown costumes and the big empty gift bags?" Bill adjusted his oversized clown wig nervously while clutching a sagging laundry sack. "Yes, plus the getaway tricycle and fake mustaches for disguise." She grinned wickedly, tapping the hand-drawn map of folding tables piled with presents. "Perfect, we crash the backyard party during cake and grab everything fast." Bill whispered, glancing at the screaming children chasing the bubble machine. "What if the kids cry or the magician tackles us again?" She rolled her eyes, smearing red nose paint across her cheek confidently. "Then we cry louder and demand extra juice boxes for pain." Bill shuddered, remembering bruises and confetti stuck in uncomfortable places. "I love stealing toys, but seven-year-olds fight really dirty." Molly snorted, hoisting the empty sack over her shoulder like Santa. "Relax, I packed glitter bombs, silly string, and total confidence." He sighed, tightening his suspenders as the birthday song started loudly next door. "Remember last time we got beaten up by a piƱata?" "That was humiliating, so tonight we steal every single present first, okay my brave friend." They moved.
created_at
"2026-09-03T08:15:34.475Z"
9/3/2026, 8:15:34 AM
model
"meta/muse-spark-1.3-20260902"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
6346
moderation_latency
(null)
generation_time
84192
tokens_prompt
75
tokens_completion
364
native_tokens_prompt
68
native_tokens_completion
4898
native_tokens_completion_images
0
native_tokens_reasoning
4571
native_tokens_cached
35
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.020863
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788423334-ysr3DdBdY8fQb2aUYsoC"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788423334-8V4lUNtBGPYAMoqYhda8"
upstream_id
"resp_6a992ca859ebd1e628634eab"
provider_responses
0
endpoint_id
"cf4f1b4e-1719-4b65-9111-7dd7635e5a2f"
id
"resp_6a992ca859ebd1e628634eab"
is_byok
false
latency
1796
model_permaslug
"meta/muse-spark-1.3-20260902"
provider_name
"Meta"
status
200
total_cost
0.020863
cache_discount
0.0000385
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 50.25%, Deviation: 0.25%
neededClean
false
wordsTotal
203
wordsDialogue
102
100.0000%