NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD6A6N56XVC4P7HGYSQ78JA
Run Details
Model:
Muse Spark 1.1 (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
19.3s
Cost:
$0.014634
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly stared at the bouncy castle that guarded the suburban backyard like an inflatable fortress. Balloons quivered in the wind while a stressed clown made balloon animals with aggressive precision. Bill adjusted his fake mustache, which was glued on upside down and shedding. Inside, past the juice boxes and screaming, sat a glorious mountain of presents wrapped in dinosaur paper, practically begging for liberation. He unfurled their master plan, a crayon blueprint stolen from the birthday boy's fridge, marked with streamers labeled lasers and toddlers labeled highly caffeinated guards. The strategy was elegant: Molly would perform a distracting magic show while Bill shoved everything into a laundry basket marked definitely not suspicious. The air smelled like frosting and extremely poor life choices. Molly sighed at the absurdity of their criminal careers. "We steal from children now for real?" Molly whispered. "Technically we steal childhood itself back," Bill replied. "That LEGO set is worth fifty bucks," Molly muttered. Bill swallowed hard. Some jobs required kneepads and dignity, but this one required protection from musical chairs, glitter, and existential judgment from five-year-olds armed with Capri Suns. They nodded solemnly, warriors about to enter battle against frosting and fury very soon tomorrow morning.
created_at
"2026-07-25T17:49:05.075Z"
7/25/2026, 5:49:05 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
17578
moderation_latency
(null)
generation_time
19211
tokens_prompt
75
tokens_completion
339
native_tokens_prompt
225
native_tokens_completion
3377
native_tokens_completion_images
0
native_tokens_reasoning
3105
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.0146335
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f7119748427d02fa4ea7"
is_byok
false
latency
356
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785001745-Yc3IVg8cdfdn7De7MFZy"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785001745-Q8adQnghdBkrMYirLRAC"
upstream_id
"resp_6a64f7119748427d02fa4ea7"
total_cost
0.0146335
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 9.90%, Deviation: 0.10%
neededClean
false
wordsTotal
202
wordsDialogue
20
100.0000%