NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD6GEW6WS08HSFBEZTKWMKM
Run Details
Model:
Muse Spark 1.1 (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
12.2s
Cost:
$0.009130
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
100.00%
Dialogue to Total Word Ratio
(details)
100.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly spread the blueprints across the kitchen table, but they were not bank vault schematics, they were a crayon map of little Timmy's backyard birthday bash, complete with a bouncy castle perimeter and a juice box station. Bill stared at the X marking the present table, guarded by what the legend labeled as extremely territorial five year olds. Molly said, "We need clown disguises and a getaway wagon." Bill winced at the memory of his last undercover birthday job, which ended with him trapped in a face painting line for three hours. The plan was simple. Infiltrate as entertainers, distract the parents with balloon animals, then vanish with a sack of hastily wrapped action figures and glittery craft kits. Bill adjusted his fake red nose nervously, already sweating. Molly said, "I will not wear squeaky shoes again ever." She had packed sensible sneakers instead, which ruined the aesthetic but improved escape velocity considerably. Bill said, "Those presents are priceless." They both knew the real score was not resale value, but finally getting revenge on Timmy's mother for never returning their casserole dish. The getaway stroller waited outside, loaded with cupcakes for bribes and courage. They nodded and marched toward destiny.
created_at
"2026-07-25T17:52:30.096Z"
7/25/2026, 5:52:30 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10959
moderation_latency
(null)
generation_time
12046
tokens_prompt
75
tokens_completion
322
native_tokens_prompt
225
native_tokens_completion
2082
native_tokens_completion_images
0
native_tokens_reasoning
1805
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.00912975
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f7dee4522b49bd5e494a"
is_byok
false
latency
428
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785001950-rUNhDioH7P6nJb162ets"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785001950-fB9lbRUM7DZBbgjmXTUU"
upstream_id
"resp_6a64f7dee4522b49bd5e494a"
total_cost
0.00912975
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
100.00%
Dialogue to Total Word Ratio
Ratio: 10.00%, Deviation: 0.00%
neededClean
false
wordsTotal
200
wordsDialogue
20
100.0000%