NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD5JQSY3PQNEBSFM2F5X4M1
Run Details
Model:
Muse Spark 1.1 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
18.9s
Cost:
$0.013873
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
99.9999%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the crayon-drawn blueprint across the getaway car's dashboard while Bill squinted skeptically. The target house pulsed with shrieking six-year-olds and a wobbling bounce castle threatening collapse. Their intel confirmed thirty juice boxes and a mountain of badly wrapped presents stacked beside the princess cake. "Bill, we need clown disguises, right?" Molly whispered. Bill tugged his striped socks higher and stared at the entrance guarded by a teenage magician. "Molly, dignity left when you rented that cupcake van." Bill muttered. The plan was simple. Infiltrate as entertainers, distract with balloon animals, then smuggle loot in empty gift bags. "Focus, it's strategic frosting camouflage." Molly hissed, dabbing sprinkles on his nose. Bill sighed, inflating a poodle that immediately popped, causing toddlers to wail louder than car alarms. Security was just one exhausted mom armed with a whistle and an impressive juice-box negotiation tactic. Molly checked her watch. The cake ceremony was their window, thirty seconds of sugar-fueled anarchy to execute the present snatch. Bill nodded solemnly, realizing they'd never stolen anything so sticky, loud, or likely to demand they stay for musical chairs afterward. They high-fived, donned squeaky shoes, and marched toward destiny, glitter, and felony-level party fouls. Finally ready.
created_at
"2026-07-25T17:36:16.245Z"
7/25/2026, 5:36:16 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
17515
moderation_latency
(null)
generation_time
18764
tokens_prompt
75
tokens_completion
343
native_tokens_prompt
225
native_tokens_completion
3198
native_tokens_completion_images
0
native_tokens_reasoning
2905
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.01387275
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f4102d492eb059bf435f"
is_byok
false
latency
719
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785000976-UnsPL8EgnoKqI1Y4AIBj"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785000976-GP22RE6jo4nJCPGNSuUQ"
upstream_id
"resp_6a64f4102d492eb059bf435f"
total_cost
0.01387275
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 9.66%, Deviation: 0.34%
neededClean
false
wordsTotal
207
wordsDialogue
20
99.9999%