NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD66G78NX95GH42902G7CCZ
Run Details
Model:
Muse Spark 1.1 (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
10.5s
Cost:
$0.007553
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.9995%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
The van smelled faintly of stolen ham sandwiches and poor life decisions. Molly spread the blueprints across the dashboard, which were actually just a crayon drawing taped to a pizza flyer. Bill nodded solemnly, adjusting his oversized clown wig that they had purchased from a discount costume bin. Their target was little Timmy Henderson's seventh birthday blowout, an event rumored to have more wrapped boxes than a department store. "Bill, are we really stealing juice boxes now?" Molly flicked a goldfish cracker off the steering wheel with the precision of a safecracker disarming lasers. According to her intel, every parent had tried to outdo each other, meaning action figures, video games, and possibly a live pony. "Molly, the real loot is that giant LEGO castle." Bill pointed at the drawing where a lopsided square was labeled DEFINITELY NOT A BOUNCY CASTLE. The plan was simple: pose as entertainers, infiltrate the backyard, and extract the presents while leading a rousing rendition of Baby Shark. Molly groaned because Bill could not clap on beat if his life depended on it. "Absolutely genius, Bill." They high-fived, knocking over a thermos, and began rehearsing balloon animals that looked suspiciously like crime confessions together always scheming.
created_at
"2026-07-25T17:47:03.794Z"
7/25/2026, 5:47:03 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
9215
moderation_latency
(null)
generation_time
10355
tokens_prompt
75
tokens_completion
321
native_tokens_prompt
225
native_tokens_completion
1711
native_tokens_completion_images
0
native_tokens_reasoning
1423
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.007553
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f6980dd836de3aa64d46"
is_byok
false
latency
264
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785001623-I4vPFKyDwIGHNsVTgbJW"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785001623-Z6AlsSsjfAVZG1IedWXE"
upstream_id
"resp_6a64f6980dd836de3aa64d46"
total_cost
0.007553
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
201
100.00%
Dialogue to Total Word Ratio
Ratio: 9.90%, Deviation: 0.10%
neededClean
false
wordsTotal
202
wordsDialogue
20
99.9995%