NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1K4VQ68H70P804FCGCB65R6
Run Details
Model:
Muse Spark 1.2 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 50% dialogue
Duration:
13.7s
Cost:
$0.014909
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 50% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly crouched behind the garden fence, peering at the bouncy castle. Bill unfolded his blueprint, which was actually a crumpled napkin. "Okay, so we go through the back gate during the piƱata?" Molly whispered. "No, we dress as clowns and walk straight in," Bill said. "Bill, you are terrified of clowns," she hissed. "I am, but kids love them! Nobody suspects a sad clown," he replied. "And how do we carry twelve presents past screaming toddlers?" she asked. "We use the present sack. Santa style," he grinned. "That's insane. The parents hired a magician," she muttered. "Perfect. I will distract him while you grab the loot," he whispered. "Fine, but if I get stuck in that bouncy castle again, you owe me," she sighed. Bill nodded solemnly as a child screamed and cake flew everywhere. The wind lifted the napkin blueprint and plastered it against Bill's face. Molly yanked it off, revealing his poorly drawn escape route through the sprinkler system. She stared, unimpressed. "We are not army crawling through sprinklers," she snapped. "We will be soaked and dignified," he protested. He adjusted his nose. She checked her watch as the birthday song began. "Ready?" he asked. "Always," she answered.
created_at
"2026-09-03T08:06:13.965Z"
9/3/2026, 8:06:13 AM
model
"meta/muse-spark-1.2-20260805"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
4557
moderation_latency
(null)
generation_time
13646
tokens_prompt
75
tokens_completion
392
native_tokens_prompt
68
native_tokens_completion
3488
native_tokens_completion_images
0
native_tokens_reasoning
3116
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.014909
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788422773-ggBXW3EoRw5U9Qxhi9mj"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788422773-p8Db9LbX2KpMLbq5Hb7x"
upstream_id
"resp_6a992a76d1f77ccaa2174330"
provider_responses
0
endpoint_id
"ffc8e76e-808a-44d2-bd21-c6ac16e721cb"
id
"resp_6a992a76d1f77ccaa2174330"
is_byok
false
latency
831
model_permaslug
"meta/muse-spark-1.2-20260805"
provider_name
"Meta"
status
200
total_cost
0.014909
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 50.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
100
100.0000%