NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD4V0592FC4ZPRWFTBPT2BT
Run Details
Model:
Muse Spark 1.1 (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
18.4s
Cost:
$0.011340
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the recycling bins, staring at the bouncy castle wheezing in Mrs. Henderson's yard. Bill taped his hand-drawn map to the fence with trembling enthusiasm. The afternoon sun made the frosting sweat and the children scream louder than any alarm system they had ever faced before. "This is insane, Bill. We're stealing from seven-year-olds?" Molly whispered. "It's not stealing, it's strategic present redistribution," Bill said, adjusting his fake mustache. Bill's mustache slipped south, revealing a ketchup stain shaped like regret. "There are literally twenty parents inside. And a clown." "Distractions, both of them. Look, Timothy gets mountains of Legos every year. He won't miss one robot dinosaur." "Two dozen robot dinosaurs. Plus gift cards." "Exactly. We go through the side gate, I distract the clown, you grab the table." "You want me to wrestle a clown while you steal Bluey wrapping paper?" "No, I wrestle the clown. You said you were great with kids." Molly sighed, watching a small army of juice boxes march past. The pile of presents glittered like shameful treasure. Somewhere, a timer ticked on the oven and on their terrible decisions today finally. "Fine. But if we get caught, we're grounded by actual moms."
created_at
"2026-07-25T17:23:18.322Z"
7/25/2026, 5:23:18 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
16926
moderation_latency
(null)
generation_time
18348
tokens_prompt
75
tokens_completion
320
native_tokens_prompt
225
native_tokens_completion
2602
native_tokens_completion_images
0
native_tokens_reasoning
2319
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.01133975
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f1069bffde8056c4470f"
is_byok
false
latency
350
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785000198-ZeLJVULkqEkXj5jajqO5"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785000198-myhMrN2l7RmHtcV7jOta"
upstream_id
"resp_6a64f1069bffde8056c4470f"
total_cost
0.01133975
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 50.25%, Deviation: 0.25%
neededClean
false
wordsTotal
203
wordsDialogue
102
100.0000%