NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD5F9X04VZJZK06CXE3WQA1
Run Details
Model:
Muse Spark 1.1 (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
11.8s
Cost:
$0.008798
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the crayon map across the hood of the getaway van, weighting corners with juice boxes. Bill squinted at the fortress of balloons and folding tables behind the suburban fence. The target was Timmy's seventh birthday, an event reportedly boasting more wrapped loot than a mall at Christmas. Their intel, a frustrated clown, had confirmed the cake would serve as distraction. Bill tapped the side entrance, near the bouncy castle, whispering that the parents would be too busy arguing about allergies to notice them in costume. Molly nodded, adjusting her oversized clown wig that smelled faintly of cheese. Their plan required impeccable timing, inflatable stealth, and zero interaction with the magician. The shadows lengthened as they finalized the approach, debating who would carry the giant sack. Bill insisted his experience with piƱatas made him qualified. Molly disagreed violently, citing his previous disaster at the puppet show heist. The plan was stupid even by their low standards but the promise of unsupervised cupcakes kept them motivated for the inevitable chase. Rain threatened overhead today softly. "We need the blue gift first, it's obviously LEGO and valuable," Bill hissed. "No way, grab everything and run fast you idiot," Molly snapped back.
created_at
"2026-07-25T17:34:23.655Z"
7/25/2026, 5:34:23 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10214
moderation_latency
(null)
generation_time
11690
tokens_prompt
75
tokens_completion
325
native_tokens_prompt
225
native_tokens_completion
2004
native_tokens_completion_images
0
native_tokens_reasoning
1738
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.00879825
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f39f506a3a07361a4835"
is_byok
false
latency
191
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785000863-suOZvKhhMDvNwNoP4yqH"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785000863-72tPwjzLSltxHGeofoy9"
upstream_id
"resp_6a64f39f506a3a07361a4835"
total_cost
0.00879825
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%