NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-03-01M1KRZ77Z3B1CXDXN914XN8ET
Run Details
Model:
Muse Glimmer 30B (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 10% dialogue
Duration:
33.9s
Cost:
$0.002183
Provider:
Phala
Evaluation summary
Score
Evaluator
Details Link
95.99%
Matches word count
(details)
36.79%
Dialogue to Total Word Ratio
(details)
66.3873%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 10% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly unfolded the map of the suburban cul-de-sac while Bill tested the squeaky rubber duck on the kitchen counter. The target was a six-year-old's birthday, a fortress of plastic tables and a piƱata shaped like a dinosaur. Presents were stacked by the door like a small mountain range, each wrapped in shiny paper that whispered value. She checked the plan taped to the fridge, complete with a timeline and a doodle of a getaway wagon. Bill adjusted his fake mustache and whispered about the noise discipline. The party would be loud enough to cover their entry through the back garden. Molly slipped on oversized clown shoes for silence, which was optimistic. She pointed at the window. Bill nodded, already rehearsing his cover story as a late delivery man. The clock ticked toward three o'clock. The plan was simple, stupid, and perfect. Bill, do you have a spare burglary glove? Yes it smells like crayons and regret. We are robbing children Bill. She tapped the plan twice and grinned. He checked his watch, then his pockets, then his conscience, which was on silent. The front door would be unguarded for exactly forty seconds.
created_at
"2026-09-03T13:57:40.231Z"
9/3/2026, 1:57:40 PM
model
"meta/muse-glimmer-30b-20260810"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
266
moderation_latency
(null)
generation_time
33737
tokens_prompt
75
tokens_completion
2060
native_tokens_prompt
118
native_tokens_completion
1980
native_tokens_completion_images
(null)
native_tokens_reasoning
1729
native_tokens_cached
117
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"stop"
service_tier
(null)
usage
0.00218298
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788443860-ih9Fxk7IJVDrUaLeiILV"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788443860-lwDiyDWC6hf97gYikRPc"
upstream_id
"req_9486e90e27489a8cbc433505f9271a8e"
provider_responses
0
endpoint_id
"6244f10e-1e45-488b-a47f-f575e0edd12b"
id
"req_9486e90e27489a8cbc433505f9271a8e"
is_byok
false
latency
224
model_permaslug
"meta/muse-glimmer-30b-20260810"
provider_name
"Phala"
status
200
total_cost
0.00218298
cache_discount
0.00003042
upstream_inference_cost
0
provider_name
"Phala"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
95.99%
Matches word count
n/a
neededClean
false
words
192
36.79%
Dialogue to Total Word Ratio
Ratio: 0.00%, Deviation: 10.00%
neededClean
false
wordsTotal
196
wordsDialogue
0
66.3873%