Skip to content
Tech Interview Prep home

Search

482 results for “AI Engineer”

Practice MCQ

A support copilot's regression eval flags 200 labelled failures: 41% invent facts the corpus never held (a price changed last week and no page carries it), 23% retrieve the right passage and then answer from memory, 19% return prose where the caller parses JSON, and 17% drop a constraint stated in the request. Pricing pages change weekly, and the caller hard-fails any response that does not validate against its schema. Which plan should the team ship?

Role practice MCQ for AI Engineer

Practice MCQ

A team claims its model is reliable enough to manage agent incident response without platform enforcement. Which counter-design should you recommend?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

A team must refresh the retrieval benchmark for a customer-facing RAG search two weeks before an embedding-model swap. They have 18 months of production query logs and two domain experts free this week; end users read only the generated answer and cannot judge retrieved passages; no incident is open, no existing sessions are pinned to the current index, and scores must stay comparable across releases. Which evaluation-set policy should they adopt?

Role practice MCQ for AI Engineer

Practice MCQ

An unattended provisioning agent caused two incidents: a run retried create_record 240 times over 25 minutes before the watchdog killed it with no result; another emitted a record id and reported success while the store held no such record, and billing closed the job. Jobs run unattended at 200/hr, tool calls take 50ms-40s, plans vary per request, and each run must stop within 3 minutes. Which termination design fixes both failures?

Role practice MCQ for AI Engineer

Practice MCQ

Complaints about confidently wrong RAG answers land 10-14 days after the call, and the call itself returned outcome: answered. The provider's request logs expire after 7 days, and the search index is republished nightly so replaying the query the next morning returns different passages. Compliance has cleared keeping call payloads with retrieved document references for 30 days. Which logging design should the team adopt?

Role practice MCQ for AI Engineer

Practice MCQ

During a threat model for risk-based approvals, reviewers identify this hazard: Approval fatigue turns dialogs into rubber stamps, while approving a vague plan allows the model to substitute different recipients or amounts afterward. Which control most directly contains it?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

The release checklist for a candidate prompt change asks only for the aggregate suite score, flat at 0.86 against the incumbent; a spot check shows billing/refunds dropped from 0.91 to 0.72, and a customer hit that case in production before release review. Product never signed off a behaviour change, the change alters model output, and all five product teams run the same property-checkable suite. What handling of prompt and model regression testing is sound?

Role practice MCQ for AI Engineer

Practice MCQ

The team must prove that embedding cache invalidation remains safe under retries, stale state, and partial failure. Which design is defensible?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

The team must prove that multi-agent liveness remains safe under retries, stale state, and partial failure. Which design is defensible?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

The team must prove that tool argument validation remains safe under retries, stale state, and partial failure. Which design is defensible?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

Three contract annotators labelled 40k support tickets with no shared guide. On a 200-item overlap check, one sends 40% of refund disputes to 'billing_general' while the others split them out. Holdout accuracy is 0.94 against the raw labels; a compliance review next week will ask what makes that figure credible, and the budget covers a second read on a sample, not a full re-label. What should the team commit to before the review?

Role practice MCQ for AI Engineer

Practice MCQ

Your agent platform is adding agent event sourcing. Which engineering decision creates the strongest operational guarantee?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

Your agent platform is adding deterministic replay. Which engineering decision creates the strongest operational guarantee?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

Your agent platform is adding idempotent tool execution. Which engineering decision creates the strongest operational guarantee?

Role practice MCQ for Agentic AI Engineer

Practice MCQ

Your agent platform is adding the GIL and agents. Which engineering decision creates the strongest operational guarantee?

Role practice MCQ for Agentic AI Engineer

Interview QA

A customer wants an AI assistant that answers questions about their internal procedures. Would you use prompting, retrieval, or fine-tuning?

Role interview QA for Forward Deployed Engineer (FDE)

Interview QA

When should a person approve an AI system's output, and how do you design that step?

Role interview QA for Forward Deployed Engineer (FDE)

Job Role

Backend Engineer

Builds and operates product services, owning APIs, data flows, reliability, and production diagnosis across databases, queues, caches, and downstream dependencies.

Job Role

Cloud Security Engineer

A Cloud Security Engineer designs and enforces controls across cloud identities, workloads, networks, data, and control planes, and proves they hold under real traffic and real incidents.

Job Role

Cybersecurity Engineer

A Cybersecurity Engineer designs and operates the controls that limit blast radius—identity, network, and data—plus the detection and response paths that catch what gets through.