NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYD6Z86F6KHNKAQECE6QQ6T3
Run Details
Model:
Muse Spark 1.1 (Reasoning, Medium)
Test:
N-Length Sentences
Scenario:
Write sentences with 20 words each
Duration:
11.1s
Cost:
$0.009080
Provider:
Meta
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
Model input
User/Human
Please write the opening scene of a murder mystery where each sentence has exactly 20 words. Please make sure that each sentence has the same length. Only return the opening itself.
Model response
Messages
Meta
AI
Detective Marlowe arrived at Blackwood Manor precisely as thunder cracked violently across the desolate, stormy midnight sky above his shoulders. The butler, pale and trembling uncontrollably, opened the heavy oak doors without uttering a single intelligible word to anyone listening. Inside the library, Lord Harrington lay sprawled upon expensive Persian carpet, blood pooling beneath his lifeless chest like spilled wine. Marlowe knelt beside the body, noting the absence of any visible weapon or forced entry nearby despite the violent struggle. Someone had meticulously arranged seven porcelain ravens around the corpse, each facing outward toward the shattered window expecting imminent dawn. From upstairs, a child's laughter echoed suddenly, though Harrington had supposedly lived completely alone for decades according to every record.
created_at
"2026-07-25T18:00:34.778Z"
7/25/2026, 6:00:34 PM
model
"meta/muse-spark-1.1-20260709"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
10368
moderation_latency
(null)
generation_time
11061
tokens_prompt
46
tokens_completion
214
native_tokens_prompt
199
native_tokens_completion
2078
native_tokens_completion_images
0
native_tokens_reasoning
1906
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"completed"
service_tier
"auto"
usage
0.00908025
router
(null)
provider_responses
0
endpoint_id
"b2b9f6f9-8880-41c1-bd0c-867650fd5238"
id
"resp_6a64f9c3fa7c482469624e72"
is_byok
false
latency
289
model_permaslug
"meta/muse-spark-1.1-20260709"
provider_name
"Meta"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1785002434-LxvTowAupBzKEBU6q3vs"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1785002434-7p93vQHwEUPYxFkkSh7p"
upstream_id
"resp_6a64f9c3fa7c482469624e72"
total_cost
0.00908025
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Meta"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
sentences
6
wordCounts
0
20
1
20
2
20
3
20
4
20
5
20