NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1K5J6HHA4CY776EPJPH3KBH
Run Details
Model:
Muse Spark 1.2 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
18.4s
Cost:
$0.019010
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly peered through the fence at the bouncing castle. Bill unfolded the party map and whispered. She nodded. He sighed. "Bill, are we seriously going to steal presents from a kids' party?" "Molly, those presents are worth a fortune, LEGO sets, bikes, gift cards, pure profit." "We are career criminals, Bill, not monsters who crash a princess birthday for loot." "Listen, the parents hired a clown, security is literally one sleepy dad with a spatula." "And the kids? They do karate chops on the cake, they are vicious." "Perfect distraction, we enter as animators, grab the present table, exit through the backyard gate." "What if they invited a magician? He will make us disappear before we get the pile." "Then we steal his rabbit too, two for one, Molly, think big." "Fine, but I wear the dinosaur costume, you get the clown nose, no arguments." "Deal, you distract the kids with roars, I will bag every wrapped box silently." "If anyone cries, we leave the juice boxes, we are thieves, not heartless." "Agreed, nobody touches the birthday cake, that frosting is sacred, even for us." "On three we go, grab, grin, and run like toddlers after sugar." "Ready, partner?" "Yes."
created_at
"2026-09-03T08:18:30.585Z"
9/3/2026, 8:18:30 AM
model
"meta/muse-spark-1.2-20260805"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
2447
moderation_latency
(null)
generation_time
18331
tokens_prompt
75
tokens_completion
311
native_tokens_prompt
68
native_tokens_completion
4462
native_tokens_completion_images
0
native_tokens_reasoning
4154
native_tokens_cached
35
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.01901
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788423510-e4RKoFusm2hdhZ1a9wDg"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788423510-xgMHYQhOvkWmizWsuXdS"
upstream_id
"resp_6a992d567e73c0f23aa04c34"
provider_responses
0
endpoint_id
"ffc8e76e-808a-44d2-bd21-c6ac16e721cb"
id
"resp_6a992d567e73c0f23aa04c34"
is_byok
false
latency
691
model_permaslug
"meta/muse-spark-1.2-20260805"
provider_name
"Meta"
status
200
total_cost
0.01901
cache_discount
0.0000385
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 90.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
180
100.0000%