NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD7NXNH4B4QGRX85E4FE2WF
Run Details
Model:
Muse Spark 1.1 (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write unattributed dialogue
Duration:
5.6s
Cost:
$0.002988
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Count dialogue tags
(details)
Model input
User/Human
You are an expert novel writer in the cozy mystery genre. Write a scene of about 200 words, with dialogue between two characters. The dialogue should be unattributed, meaning the speaker is not explicitly mentioned. The scene should be set in a small town bakery. Only return the story, nothing else.
Model response
Messages
Meta
AI
The bell over Willow Lane Bakery jingled, letting in a gust of cold November air and the smell of wet leaves. The display case was already half-empty, only a few scones and one last apple crumble remaining. Behind the counter, flour dusted every surface, and a chalkboard read: CLOSED FOR PRIVATE TASTING - Noon. "You're early. Tasting isn't for another hour, and you weren't invited." "Neither was half the town, but we all heard you finally perfected Mrs. Abernathy's lemon tart recipe. The one she took to her grave." "That recipe died with her, and you know it. I found this in my grandmother's recipe box. Completely different." "Really? Because Mrs. Abernathy's niece says you were in her kitchen the day she passed, measuring sugar while the poor woman was measuring her last breaths." "She offered me tea! I offered to help clean up. That's what neighbors do here." "Neighbors don't usually pocket handwritten recipe cards. Or leave floury fingerprints on the back door." "Are you here to buy something or just stir up trouble like your newspaper usually does?" "Both. I'll take the apple crumble. And the truth, if you have a slice of that left."
created_at
"2026-07-25T18:12:57.655Z"
7/25/2026, 6:12:57 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
3497
moderation_latency
(null)
generation_time
5466
tokens_prompt
78
tokens_completion
296
native_tokens_prompt
225
native_tokens_completion
637
native_tokens_completion_images
0
native_tokens_reasoning
363
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.0029885
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64fcaab551e9f2e59845bb"
is_byok
false
latency
387
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785003177-iTfuK6eGtBZncWwJZzhK"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785003177-l7S3M9iFrBhp9Pxr0rph"
upstream_id
"resp_6a64fcaab551e9f2e59845bb"
total_cost
0.0029885
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Count dialogue tags
n/a
neededClean
false
dialogueTags
(empty)