NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD4RMYQ38NQQV4G3E5KZ1BY
Run Details
Model:
Muse Spark 1.1 (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
12.4s
Cost:
$0.008048
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
93.47%
Dialogue to Total Word Ratio
(details)
96.7341%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly and Bill crouched behind the bounce house, whispering fiercely. "Molly, this is a new low. Stealing from eight-year-olds? Really?" "Bill, it's not stealing, it's strategic present redistribution. Look at that pile." "That's like fifty Bluey wrapped boxes. You want us to get beaten up by moms?" "The moms are distracted by mimosas. We go in as clowns. In and out." "We don't have clown costumes, Molly. We have ski masks." "Same vibe. Kids love terrifying clowns. You distract the magician, I grab the loot." "You want me to fight a magician? What if he actually makes me disappear?" "He does balloon animals, Bill, not actual magic. Grow up." "Okay, okay. What's our escape vehicle?" "The little train that goes choo-choo around the yard. It's our getaway train." "Genius. Absolutely genius. Let's get those LEGOs." "Wait, do we split fifty-fifty?" "No, you get forty, I get sixty, I planned this brilliant operation." "That pile has a Barbie Dreamhouse, Molly. I want the Dreamhouse." "Fine, you can have the Dreamhouse if I can have the Nintendo Switch." "Deal. But if we get caught, we pretend we are the entertainment." "Agreed. Now put on this ridiculous red nose." "Let's steal some joy."
created_at
"2026-07-25T17:22:01.311Z"
7/25/2026, 5:22:01 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10960
moderation_latency
(null)
generation_time
12244
tokens_prompt
75
tokens_completion
306
native_tokens_prompt
225
native_tokens_completion
1844
native_tokens_completion_images
0
native_tokens_reasoning
1542
native_tokens_cached
64
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.00804785
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f0b9a477e1ea845b4c32"
is_byok
false
latency
606
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785000121-qEq1TmZuOXMsAz5BYO9j"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785000121-xAJXHzaM9FXK6fQvocy6"
upstream_id
"resp_6a64f0b9a477e1ea845b4c32"
total_cost
0.00804785
cache_discount
0.0000704
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
93.47%
Dialogue to Total Word Ratio
Ratio: 95.10%, Deviation: 5.10%
neededClean
false
wordsTotal
204
wordsDialogue
194
96.7341%