AI Engineering Learning Paths
Choose a core domain to build depth, or a career route that sequences the same lessons around the work you want to become capable of doing.
Four core paths. One discipline.
Each domain opens a guided lesson sequence and the capabilities it develops. Every capability links to the closest practical lesson.
Compare six career routesChoose by the work, not the title.
Titles vary between teams. Start with the problems you want to own, then inspect the responsibilities, baseline, evidence, and gaps before choosing specialist lessons.
-
01
Engineering foundations
Full-stack boundaries, data, architecture, reliability, security, and production operations.
-
02
AI application foundations
Model interfaces, grounding, evaluation, production behavior, and operations.
-
03
Specialist practice
Choose a work family, close its baseline gaps, and produce role-shaped evidence.
Which work would you want to repeat every week?
01 Work familyCustomer AI DeploymentForward-Deployed AI Engineer · Field AI Engineer · AI Solutions Engineer
Turn a real customer workflow into a small, measurable AI system, then stay close enough to the rollout to learn where it breaks.
- Guided route
- 12 specialist lessons · 865 minutes
- Baseline
- Software delivery plus AI application fundamentals
What you would own
- Observe the workflow, users, exceptions, and hidden handoffs.
- Reduce the request to the smallest useful end-to-end slice.
- Integrate grounding, evaluation, and production controls.
- Run a measured pilot and turn feedback into the next system change.
Fit and boundary
Good fit if: you like ambiguous user problems, fast technical iteration, and shared ownership after launch.
Boundary: this is not sales engineering or generic consulting. The proof is a working, measured system that you can operate.
Portfolio proof
Ship a workflow dossier, a grounded prototype, an evaluation set, and a pilot plan as one evidence bundle.
- Named assumptions and the riskiest test
- Measured task quality and failure cases
- Rollout, rollback, and feedback ownership
Course coverage and gaps
The route covers discovery, risk, RAG, evaluation, production, metrics, rollout, and feedback.
Still earned elsewhere: customer domain expertise, stakeholder trust, procurement constraints, and ownership under live production pressure.
02 Work familyDeveloper Experience and EducationAI Developer Relations Engineer · AI Developer Advocate · Developer Experience Engineer
Make an AI capability understandable, runnable, and trustworthy for developers, then feed their friction back into the product.
- Guided route
- 11 specialist lessons · 905 minutes
- Baseline
- Software fundamentals, API use, and clear technical writing
What you would own
- Build integrations and examples that survive a clean setup.
- Explain API, tool, protocol, and skill contracts precisely.
- Reproduce developer friction instead of guessing at it.
- Turn support signals into documentation, tooling, and product feedback.
Fit and boundary
Good fit if: you enjoy building, teaching, debugging with other developers, and making difficult systems legible.
Boundary: this is not content-only marketing. Credibility comes from runnable technical work and accurate explanations.
Portfolio proof
Publish a developer onboarding package with a working integration, examples, a reusable agent package, and a friction report.
- Fresh-environment setup evidence
- Positive, negative, and failure examples
- Feedback linked to a concrete improvement
Course coverage and gaps
The route covers APIs, tool contracts, MCP, Agent Skills, packaging, evaluation, and feedback.
Still earned elsewhere: live audience practice, community judgment, adoption analytics, editorial depth, and sustained developer support.
03 Work familyAI Data SystemsAI Data Engineer · Machine Learning Data Engineer · Retrieval Engineer
Build the data and retrieval pipelines that let training, evaluation, and production AI behavior use trustworthy evidence.
- Guided route
- 11 specialist lessons · 915 minutes
- Baseline
- Python, data structures, statistics, and pipeline fundamentals
What you would own
- Ingest, transform, version, and validate training or retrieval data.
- Build embedding, indexing, retrieval, and evaluation pipelines.
- Define data quality checks and investigate silent drift.
- Expose lineage, freshness, cost, and runtime health.
Fit and boundary
Good fit if: you enjoy pipelines, data quality, reproducibility, and debugging systems that fail far from the user interface.
Boundary: this route focuses on AI data products. It does not replace the broader warehouse, database, and platform depth of data engineering.
Portfolio proof
Ship a versioned document-to-retrieval pipeline with quality gates, evaluation data, and an operational report.
- Reproducible ingestion and lineage
- Retrieval quality and freshness measures
- Failure recovery and observability evidence
Course coverage and gaps
The route covers data management, features, pipelines, embeddings, context, RAG, evaluation, production, and observability.
Still earned elsewhere: advanced SQL, warehouse architecture, governance, privacy operations, and large-scale distributed data systems.
04 Work familyAgent Systems EngineeringAgent Systems Engineer · Agentic AI Engineer · AI Agent Engineer
Engineer the runtime around a tool-using model so context, memory, authority, orchestration, failure, and evidence remain explicit.
- Guided route
- 14 specialist lessons · 865 minutes
- Baseline
- LLM application foundations plus typed tool interfaces
What you would own
- Design tool contracts and the observe, decide, act loop.
- Control context, memory, state, and durable execution.
- Choose orchestration boundaries and termination policy.
- Threat-model authority and evaluate complete trajectories.
- Operate the runtime with traces and explicit failure controls.
Fit and boundary
Good fit if: you enjoy runtime design, state machines, distributed coordination, safety boundaries, and difficult failure analysis.
Boundary: this is systems engineering around model behavior, not a promise that adding an agent loop makes a product autonomous.
Portfolio proof
Ship a bounded tool-using runtime with memory, orchestration, a threat model, trajectory evals, and a failure runbook.
- Deterministic tool and state traces
- Permission, sandbox, and injection controls
- Termination, recovery, and evaluation evidence
Course coverage and gaps
The route covers tools, MCP, loops, context, memory, graphs, orchestration, security, evaluation, runtimes, and observability.
Still earned elsewhere: provider-specific infrastructure, high-scale distributed operation, latency engineering, and production ownership with a team.
05 Work familyLLM Product EngineeringApplied AI Engineer · LLM Engineer · AI Product Engineer
Turn model capability into useful product behavior that is grounded, evaluated, guarded, cost-aware, and recoverable in production.
- Guided route
- 12 specialist lessons · 885 minutes
- Baseline
- Software engineering plus LLM foundations
What you would own
- Design model-facing interfaces and structured contracts.
- Ground behavior with context, retrieval, and tools.
- Build task evaluations before optimizing the feature.
- Control safety, cost, latency, caching, and fallbacks.
- Release the complete feature with observable behavior.
Fit and boundary
Good fit if: you want to connect product needs to model behavior and own the software around the model.
Boundary: this is not foundation-model research or model training. The work begins where a model capability meets a real product constraint.
Portfolio proof
Ship a grounded product feature with structured output, tools, an eval set, cost and latency budgets, and a guarded release.
- Representative success and failure cases
- Quality, cost, and latency tradeoffs
- Fallback, release, and rollback evidence
Course coverage and gaps
The route covers prompting, structured output, embeddings, context, RAG, tools, evaluation, cost, guardrails, production, gateways, and release.
Still earned elsewhere: product discovery, interaction design, real user research, domain regulation, and operating a feature under sustained traffic.
06 Work familyAI Evaluation and ReliabilityAI Evaluation Engineer · AI Reliability Engineer · Machine Learning Site Reliability Engineer
Make model and agent behavior measurable, expose failure before release, and build operational controls for what still fails in production.
- Guided route
- 12 specialist lessons · 750 minutes
- Baseline
- Statistics, software testing, and production systems
What you would own
- Define evaluation sets, metrics, graders, and failure taxonomies.
- Instrument model, agent, and serving behavior.
- Build release gates, experiments, and regression detection.
- Test load, degradation, recovery, and incident response.
- Connect evidence to rollout and operational decisions.
Fit and boundary
Good fit if: you enjoy statistics, adversarial testing, observability, release judgment, and learning from incidents.
Boundary: this is broader than offline model accuracy. Reliability includes the application, runtime, infrastructure, and response process.
Portfolio proof
Ship a behavioral evaluation harness connected to traces, a release gate, a load or failure experiment, and an incident runbook.
- Versioned cases and metric rationale
- Regression and rollout decisions
- Observed recovery and residual risk
Course coverage and gaps
The route covers model, LLM, and agent evaluation, observability, serving metrics, experiments, load, canary release, chaos, and SRE.
Still earned elsewhere: real on-call experience, organization-specific incident process, production traffic, compliance evidence, and cross-team release authority.
Building and Deploying AI Applications
Move from the first model-facing interface to grounded behavior, evaluation, safeguards, and production operation. The application is the whole system around the model.
Turn intent into a bounded request with explicit inputs, outputs, and failure behavior.
Open lesson 02 · contractsStructured Generation ContractsMake generated data parseable, validated, and safe to pass into application code.
Open lesson 03 · groundingEvidence RepresentationRepresent, retrieve, and place evidence where the model can use it.
Open lesson 04 · retrievalRetrieval and FreshnessBuild the ingestion, search, ranking, citation, and freshness loop around generation.
Open lesson 05 · evidenceBehavioral Evaluation GatesDefine acceptable behavior, collect cases, score outcomes, and gate regressions.
Open lesson 06 · operationServing and RecoveryServe, observe, release, recover, and control cost under real traffic.
Open lessonSoftware Engineering Fundamentals
Coding agents reduce typing, not engineering judgment. Learn to steer tradeoffs across the application stack, data, architecture, security, reliability, and production operations.
Connect request handling, streaming, persistence, fallbacks, health checks, and deployment into one working system.
Open representative lesson 02 · dataData Lifecycle and StorageChoose representations, validation, versioning, retention, and freshness from the access patterns the application needs.
Open representative lesson 03 · architectureSystem Architecture and BoundariesDesign one explicit boundary: inputs, outputs, errors, permissions, and state before a capability enters a larger system.
Open representative lesson 04 · assuranceSecure and Resilient SystemsAudit secrets, permissions, dependencies, data handling, and release evidence before production.
Open representative lesson 05 · operationProduction Scale and Service OwnershipDefine service objectives, watch health signals, prepare runbooks, and practice evidence-based incident response.
Open representative lessonAgent-Assisted Engineering
A coding agent is useful when the task, context, tools, feedback, and stop condition form a dependable harness. Learn to shape that system around real repository work.
Turn a request into scope, constraints, permissions, evidence, and a stopping rule.
Open lesson 02 · planEvidence-Based PlanningInspect the repository before proposing the smallest coherent change.
Open lesson 03 · harnessAgent WorkbenchEngineer the loop, context boundary, tools, transcript, and termination policy.
Open lesson 04 · contextInstructions and MemoryPlace durable guidance at the correct scope and keep runtime state observable.
Open lesson 05 · feedbackRuntime FeedbackFeed compiler, test, browser, and wire evidence back into the next decision.
Open lesson 06 · verifyVerification and ReviewProve the requested behavior independently of the agent's own completion claim.
Open lesson 07 · delegateIsolated DelegationSplit bounded work across agents without sharing ambiguous ownership or state.
Open lesson 08 · improveDurable ImprovementConvert corrections into tests, instructions, tooling, and reusable constraints.
Open lessonProduct Judgment and Delivery
Before implementation, decide what outcome matters, what evidence supports the work, which risk deserves attention, and how you will know the change helped.
Define the changed state you want before discussing features or implementation.
Open lesson 02 · observeWorkflow DiscoveryStudy how the work happens now, including exceptions, handoffs, and hidden labor.
Open lesson 03 · riskAssumptions and RiskExpose what must be true and test the uncertainty that could invalidate the build.
Open lesson 04 · sliceTestable SlicesChoose the smallest end-to-end change that can produce decision-quality evidence.
Open lesson 05 · specifyExecutable SpecificationsMake constraints and acceptance observable without removing implementation judgment.
Open lesson 06 · measureSuccess MetricsConnect product outcomes to leading, guardrail, and operational measures.
Open lesson 07 · stageRelease StrategyMatch prototype, pilot, or production investment to the evidence you need next.
Open lesson 08 · ownFeedback OwnershipAssign who reads the signal, makes the decision, and changes the system.
Open lesson