The search giant's AI-assisted coding rounds signal something deeper than a policy update. They expose a truth the industry has long resisted: the best AI user in the room is always the most knowledgeable one.
For decades, the coding interview was tech's most sacred ritual. Whiteboard in hand, candidate across the table — no hints, no documentation, no tools. Just raw recall versus a ticking clock. Google perfected this format, and the industry genuflected accordingly. What Google tested, the rest of Silicon Valley tested.
That era just ended.
According to an internal document reviewed by Business Insider, Google is piloting a new interview format that allows — and actively evaluates — the use of AI during a live coding session. Starting in the second half of 2026, select junior and mid-level software engineering candidates will be permitted to use Gemini, Google's own AI assistant, during the code comprehension round. Instead of writing from scratch, candidates will read, debug, and optimise an existing codebase alongside an AI assistant.
The rationale is almost disarming in its honesty: three-quarters of new code written inside Google is now AI-generated. The company sees little logic in testing candidates on a workflow that no longer reflects how engineers actually work.
75%
of new Google code is now AI-generated
22%
of job seekers already use AI in live interviews
4+
major tech firms now allow AI in coding rounds
"I guess this is like asking a kid to take a math test without a calculator."
— Emily Cohen, Head of People & Operations, Cognition AI
What Google Is Actually Testing
Read the fine print and something important surfaces. Interviewers will explicitly evaluate "AI fluency" — the ability to engineer effective prompts, validate AI output, and debug when the model gets it wrong. This is not a rubber stamp on AI dependency. It is a structured test of how well a candidate commands AI to reach a correct outcome.
That distinction matters enormously. Because the moment you allow AI into an interview room, you introduce a variable that separates candidates faster than any whiteboard ever could: domain depth.
Editor's Analysis
Google is not lowering the bar. It is raising it in a direction most candidates are not prepared for. Knowing how to open a chat interface means nothing. Knowing how to ask the right question, recognise a wrong answer, and push toward the optimal solution — that requires genuine expertise.
The Hidden Thesis: Subject Matter Experts Always Win
When two candidates sit down with the same AI tool and face the same problem, what separates them is not who can type faster. It is who understands the problem deeply enough to know whether the AI's answer is good, mediocre, or dangerously wrong.
AI is a multiplier. Like any multiplier, it amplifies what you bring to it. A shallow prompt from a novice returns shallow output — faster. A precise, expert prompt returns insight the novice would not even know to ask for. The AI does not close the expertise gap. It widens it.
The clearest way to see this is to compare how an expert and an average user approach the same task across different fields.
— · —
Expert vs. Average: The Same AI, Very Different Results
Finance & Investment Analysis
CFA / Finance Expert
"Analyse this company's free cash flow trend over 5 years, flag divergence from reported net income, and identify non-cash adjustments that could indicate earnings quality issues. Use DuPont decomposition to isolate the ROE drivers."
Applies a named framework, anticipates where models mislead, and scopes the request precisely.
Average Joe
"Is this a good stock to buy?"
No framework, no context. The AI returns generic caveats and a surface-level summary useful to no one.
Outcome
Expert gets →A structured earnings quality report with flagged ratios, a DuPont breakdown, and actionable investment insight.
Average Joe gets →"This stock has both risks and opportunities." A disclaimer-laden non-answer.
Software Engineering
Senior SWE
"This Node.js service handles 10k req/s. The p99 latency spikes every ~4 minutes. I suspect GC pressure from large object allocations in the request pipeline. Suggest targeted profiling steps and refactor options that avoid heap fragmentation."
Diagnoses before asking. Provides system context, names the likely root cause, requests a scoped solution.
Junior / Non-Expert
"My app is slow. How do I fix it?"
No context, no hypothesis. The AI returns a generic checklist — none of which may apply to this architecture.
Outcome
Expert gets →Targeted profiling commands, specific refactoring patterns for the GC issue, and benchmark strategies tailored to their stack.
Average Joe gets →A 10-point generic optimisation article. Hours of irrelevant debugging ahead.
Medicine & Clinical Decision-Making
Attending Physician
"Patient is a 58-year-old male, T2DM, CKD stage 3, new onset atrial fibrillation. Current meds: metformin, lisinopril. Evaluate anticoagulation options factoring renal dosing constraints and bleeding risk using CHA₂DS₂-VASc and HAS-BLED."
Applies structured clinical scoring tools, accounts for comorbidities, requests pharmacokinetically appropriate options — not a general list.
Average Patient
"My heart is beating weird. What medicine should I take?"
Cannot specify the condition, evaluate the output, or safely act on the answer. The AI rightly hedges.
Outcome
Expert gets →A nuanced comparison of apixaban vs. rivaroxaban with renal-adjusted dosing and scored bleeding risk — actionable clinical insight.
Average Joe gets →"Please consult your healthcare provider." A dead end.
Legal & Contract Analysis
Corporate Attorney
"Review this SaaS MSA for one-sided indemnification clauses, uncapped liability exposure, and auto-renewal terms conflicting with enterprise procurement policies. Flag any IP ownership ambiguity in the work-for-hire clause."
Uses precise legal terminology, defines exact scope of review, knows which clauses carry real risk.
Average Business Owner
"Can you check this contract and tell me if it's okay to sign?"
No legal framework applied. The AI may flag obvious issues but misses the subtle risk clauses an expert catches immediately.
Outcome
Expert gets →A clause-by-clause risk register with negotiation leverage points and recommended redlines — boardroom-ready.
Average Joe gets →"This contract seems standard, but consult a lawyer for anything binding." Signing blind.
Education & Curriculum Design
Curriculum Specialist
"Design a 3-lesson sequence for 2nd graders on place value using Bruner's CPA (Concrete–Pictorial–Abstract) progression. Include formative assessment checkpoints and differentiation strategies for students reading 1–2 grade levels below."
Applies a named pedagogical framework, specifies learner profile, structures output to match how effective instruction is actually designed.
Non-Educator Parent
"How do I teach my kid about numbers?"
Too broad, no grade anchor, no framework. The AI returns generic activities disconnected from how the child's classroom teaches the concept.
Outcome
Expert gets →A structured lesson sequence with scaffolded objectives, assessment rubrics, and differentiated extensions — ready to deploy.
Average Joe gets →"Try counting blocks together!" Engaging, but no pedagogical depth or progression.
— · —
"The AI does not close the expertise gap. It widens it. Every field, every task, every prompt."
— TIBLOGICS AI Times
The Implication for Your Career
Across every field above, the pattern is identical. The expert does not just get a better answer — they get an answer that is actually usable. The average user gets noise they cannot evaluate. And that gap compounds over time: the expert uses AI to accelerate their expertise, while the novice uses AI to bypass learning they have not yet done — and eventually stalls when the outputs stop being good enough.
Google's pilot encodes this reality into hiring. By evaluating prompt engineering, output validation, and debugging during a live session, they are measuring something specific: can you tell when the AI is wrong? That question has only one honest answer — not unless you know the subject.
The TIBLOGICS Perspective
The professionals most at risk in the AI era are not those who lack AI skills. They are those who lack domain skills and expect AI to cover the gap. It never does. Not sustainably. Go deeper in your field. The best prompt engineer in any room is always the person who knows the subject so well they can tell when the model is lying.
A Broader Industry Reckoning
Google is not moving alone. Meta launched AI-enabled coding rounds in late 2025. Canva publicly stated it expects engineering candidates to use Copilot, Cursor, and Claude during technical interviews. Shopify and Rippling followed. The whiteboard-only interview is an artefact of a pre-AI world — and the industry knows it.
What remains to be seen is whether companies will follow Google's lead in structuring the evaluation — not just opening the door to AI, but actively measuring how candidates interact with it. Allowing AI is easy. Building an assessment framework that distinguishes an expert using AI from a novice leaning on it — that requires deep thinking about what professional competence actually means in 2026.
Google is betting on expertise. The smartest companies always were.
Bonus Take
From the Founder's Desk — Tieyiwe Bassole
Tieyiwe Bassole · Founder, TIBLOGICS · AI Implementation Strategist
I want to add something the industry has not yet talked about — and I think it is going to become a very real evaluation metric sooner than people expect.
Token consumption.
Think about it. If companies are already allowing AI in interviews, the next logical step is instrumenting the session. And once you can measure how a candidate uses an AI tool in real time, one of the most revealing signals you could capture is: how many tokens did it take them to get to the right answer?
"The expert arrives at the solution in three precise prompts. The novice burns twenty trying to figure out what question to ask."
— Tieyiwe Bassole, Founder, TIBLOGICS
Here is my thesis: experts will consistently reach the correct solution while consuming significantly fewer tokens than novices. Not because they type less — but because they operate from a much shorter learning curve. The novice walks into the session still figuring out the domain and the tool simultaneously. Their token spend looks like exploration. The expert already knows what they are looking for. Their token spend looks like execution.
An expert software engineer does not ask the AI to explain what garbage collection is before asking about GC pressure. A seasoned finance analyst does not prompt the AI to define free cash flow before requesting a DuPont decomposition. They skip the orientation phase entirely — and that compression shows up directly in the token log.
Illustrative token spend to reach correct solution — same task, same AI
Domain expert
~1,200 tokens
Intermediate user
~3,800 tokens
Novice / no domain
~7,500 tokens
This matters far beyond interviews. In a real work environment, token spend is operating cost. A team of shallow AI users burning 6× more tokens than a team of domain experts to produce equivalent output is a direct line to eroded margins and slower delivery. Token efficiency is about to become a proxy for professional competence — and at scale, a real line item on the P&L.
The Prediction
Within the next two hiring cycles, at least one major tech company will add token efficiency to their AI interview rubric — measuring not just whether the candidate found the solution, but how economically they got there. Screenshot this.
Google opened the door to AI in interviews. The next evolution is measuring the quality of how candidates use it. And when that happens, the token log will tell you everything about who actually knows their craft — and who was hoping the AI would figure it out for them.
Expertise was never optional. It just became measurable in a brand new way.
TIBLOGICS AI Times · tiblogics.com · © 2026
AI Hiring
Google
Future of Work
Expertise
Token Efficiency
Prompt Engineering