NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCSKDCRQXCZP2FQ3WHW8T64
Run Details
Model:
Gemini 3.5 Flash Lite (Reasoning)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
1.7s
Cost:
$0.000475
Provider:
Google AI Studio
Evaluation summary
Score
Evaluator
Details Link
0.00%
Matches word count
(details)
0.00%
Dialogue to Total Word Ratio
(details)
0.0000%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly whispered, Are you certain the toddler security force is asleep during naptime? Bill adjusted his neon green ski mask and replied, Trust me, a sugar crash is more predictable than a laser grid. But what about the piƱata guard dog? It is a Poodle named Princess, Molly, we can outrun pastry frosting. Do you have the burlap sacks for the loot? Right here, oversized and monogrammed for maximum criminal elegance. Excellent, because little Timmy is getting a minimalist birthday this year. Are we stealing the cake too? Obviously, frosting is currency on the black market. Just follow my lead past the bouncy castle. If we get captured by six-year-olds, I am blaming your terrible getaway plan. There is no plan, only chaos and stolen Lego sets. Then let us go ruin some childhood memories. Hand me the glitter glue, we might need a diversion.
created_at
"2026-07-25T14:06:55.395Z"
7/25/2026, 2:06:55 PM
model
"google/gemini-3.5-flash-lite-20260721"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
271
moderation_latency
(null)
generation_time
1638
tokens_prompt
75
tokens_completion
213
native_tokens_prompt
67
native_tokens_completion
182
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"STOP"
service_tier
"default"
usage
0.0004751
router
(null)
provider_responses
0
endpoint_id
"6bd8f433-79e1-416c-b407-1772eb796c9a"
id
"_8Jkau31HauX-8YPndmxsAk"
is_byok
false
latency
271
model_permaslug
"google/gemini-3.5-flash-lite-20260721"
provider_name
"Google AI Studio"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784988415-oh4xJY4vjEmnG98Gn2mO"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784988415-7hwpdilzoQ5WqTS2Wuar"
upstream_id
"_8Jkau31HauX-8YPndmxsAk"
total_cost
0.0004751
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Google AI Studio"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
0.00%
Matches word count
n/a
neededClean
false
words
145
0.00%
Dialogue to Total Word Ratio
Ratio: 0.00%, Deviation: 90.00%
neededClean
false
wordsTotal
147
wordsDialogue
0
0.0000%