NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-09-02-01M1HGBYEY39SW2DMBXXR43VMT
Run Details
Model:
Gemini 3.8 Flash (Reasoning, Medium)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
10.1s
Cost:
$0.013772
Provider:
Google AI Studio
Evaluation summary
Score
Evaluator
Details Link
100.00%
Matches word count
(details)
97.17%
Dialogue to Total Word Ratio
(details)
98.5831%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
"The primary target is Timmy's bouncy castle," Molly whispered, tapping her crayon map. "What kind of security are we facing?" Bill asked. "Two mothers, high on chardonnay, and a clown named Bongo who clearly did time in Sing Sing." "Brutal. What about the perimeter?" "Defended by hostile toddlers wielding sticky apple juice boxes. One touch ruins your suede jacket." "Not the suede! How do we extract the loot?" "The present pile sits behind the chocolate fountain. We deploy the diversion at three sharp." "And what is this grand diversion?" "You dress as Peppa Pig and fake an ankle sprain. While they gawk, I bag the Legos, the Nintendo, and the remote-control helicopter." "Why do I always wear the pig suit?" Bill grumbled. "Because you have the snout for it, darling. Listen, the pinata bursts first. When the sugar-crazed demons scramble for candy, that is our window." "Are we really robbing a six-year-old?" "Timmy has an offshore piggy bank and an attitude problem. He called my getaway Prius a golf cart." "That monster. Hand over the snout." "Remember, silent sneakers. Step on a squeaky toy, and we are swarmed." "I was born ready. For the Legos!" "For the Legos," she agreed.
created_at
"2026-09-02T16:48:51.175Z"
9/2/2026, 4:48:51 PM
model
"google/gemini-3.8-flash-20260902"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
1114
moderation_latency
(null)
generation_time
10029
tokens_prompt
75
tokens_completion
1343
native_tokens_prompt
67
native_tokens_completion
3659
native_tokens_completion_images
0
native_tokens_reasoning
3355
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"STOP"
service_tier
"default"
usage
0.0137715
router
(null)
user_agent
"langchainjs-openai/1.0.0 ((node/v24.19.0; linux; x64))"
http_referer
(null)
request_id
"req-1788367731-MojjmHaB8AfHmKPRRz1J"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1788367731-TkSg7cdifiwSuSgOG3oY"
upstream_id
"c1OYatHRDpCS_uMPgYu7qQE"
provider_responses
0
endpoint_id
"d376f781-516d-42f9-936d-46f19abca6cc"
id
"c1OYatHRDpCS_uMPgYu7qQE"
is_byok
false
latency
1114
model_permaslug
"google/gemini-3.8-flash-20260902"
provider_name
"Google AI Studio"
status
200
total_cost
0.0137715
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Google AI Studio"
response_cache_source_id
(null)
data_region
"global"
workspace_id
"97e315e5-d303-487d-83c1-83180e8a13d4"
Evaluation details
Result
Evaluator
Details
Meta Data
100.00%
Matches word count
n/a
neededClean
false
words
200
97.17%
Dialogue to Total Word Ratio
Ratio: 94.12%, Deviation: 4.12%
neededClean
false
wordsTotal
204
wordsDialogue
192
98.5831%