NC Bench
Overview
About
Tests
Categories
Models
Vendors
Compare
Run: 2026-07-25-01KYCTP777Z4ZQN41E77A7ZS01
Run Details
Model:
Gemini 3.5 Flash Lite (Reasoning, Minimal)
Test:
Dialogue tags
Scenario:
Write 200 words with 90% dialogue
Duration:
1.8s
Cost:
$0.000748
Provider:
Google AI Studio
Evaluation summary
Score
Evaluator
Details Link
0.00%
Matches word count
(details)
85.07%
Dialogue to Total Word Ratio
(details)
42.5341%
Overall
Model input
User/Human
You are a author of comedic heist books. Please write a 200 word scene that has 90% words of dialogue between two thieves named Molly and Bill who are planning to break into a children's party to steal all the presents. Only include the prose of the scene without any markdown or commentary.
Model response
Messages
Meta
AI
Molly whispered, "We strike at three." Bill checked his stopwatch. "During pinata annihilation?" "Exactly. Total chaos." "What about security?" "A hyperactive golden retriever named Buster." "Lethal?" "Only if you're allergic to sloppy kisses." "Got the gear?" "Grappling hooks, night-vision goggles, and industrial trash bags." "Why trash bags?" "For maximum loot efficiency, Bill." "Should we wear all-black?" "Obviously. We are elite operatives." "Of petty larceny." "Artisans of acquisition. Now, review the target." "Master Lucas turns seven. His parents bought him the deluxe Lego Death Star." "I am going to crush it." "With your bare hands?" "With absolute glee. And maybe a crowbar." "What about the juice boxes?" "Leave the juice boxes." "Are you sure? High resale value on the black market." "We are master criminals, not monsters." "Right. Leave the juice boxes. Take the ponies." "Ponies aren't real, Bill." "The little plastic ones, Molly." "Oh. Right. Grab the purple one with the sparkles." "Naturally. Are we taking cake?" "Only if it's chocolate." "Obviously."
created_at
"2026-07-25T14:25:55.952Z"
7/25/2026, 2:25:55 PM
model
"google/gemini-3.5-flash-lite-20260721"
app_id
182717
external_user
(null)
streamed
true
cancelled
false
latency
306
moderation_latency
(null)
generation_time
1693
tokens_prompt
75
tokens_completion
292
native_tokens_prompt
67
native_tokens_completion
291
native_tokens_completion_images
0
native_tokens_reasoning
0
native_tokens_cached
0
num_media_prompt
(null)
num_input_audio_prompt
(null)
num_media_completion
0
num_search_results
(null)
num_fetches
(null)
web_search_engine
(null)
origin
"https://ncbench.com/"
is_byok
false
finish_reason
"stop"
native_finish_reason
"STOP"
service_tier
"default"
usage
0.0007476
router
(null)
provider_responses
0
endpoint_id
"6bd8f433-79e1-416c-b407-1772eb796c9a"
id
"dMdkaoBJ_pr-4w-vz7yxDg"
is_byok
false
latency
306
model_permaslug
"google/gemini-3.5-flash-lite-20260721"
provider_name
"Google AI Studio"
status
200
user_agent
"langchainjs-openai/1.0.0 ((node/v24.18.0; linux; x64))"
http_referer
(null)
request_id
"req-1784989555-31INvVfY7LrOYtIV01VW"
session_id
(null)
preset_id
(null)
api_type
"completions"
id
"gen-1784989555-94DGarvZfnleUtl7rDil"
upstream_id
"dMdkaoBJ_pr-4w-vz7yxDg"
total_cost
0.0007476
cache_discount
(null)
upstream_inference_cost
0
provider_name
"Google AI Studio"
response_cache_source_id
(null)
data_region
"global"
Evaluation details
Result
Evaluator
Details
Meta Data
0.00%
Matches word count
n/a
neededClean
false
words
162
85.07%
Dialogue to Total Word Ratio
Ratio: 96.34%, Deviation: 6.34%
neededClean
false
wordsTotal
164
wordsDialogue
158
42.5341%