Console Logo

Console

Research Engineer

Posted 22 Days Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
200K-350K Annually
Mid level
In-Office
San Francisco, CA, USA
200K-350K Annually
Mid level
Build systems that measure, debug, and improve autonomous agents in production: eval and optimization loops, offline replay and tracing, labeling workflows, prompt/program optimization, and fine-tuning/adaptation of specialist models to improve reliability, latency, and quality across enterprise contexts.
The summary above was generated by AI

Console is building the AI agents that power autonomous enterprises. We use AI to autonomously resolve IT, HR, Legal, Finance, Security, and Ops tasks the moment they come in from Slack or Teams. We're building toward a future where enterprise operations are fully autonomous; with no tickets, approvals, or manual processes slowing companies down.

Our platform enables companies like Databricks, Figma, and Cursor to delegate entire categories of work to AI agents that understands their org, takes action across systems, and gets smarter over time. We've raised $29M, backed by Thrive Capital (investors in OpenAI, Stripe, GitHub) and a small group of founders who have built category-defining companies.

We’re unusually selective about the people we work with. We hire builders who span functions, optimize for speed, and care deeply about their craft. We do not care about pedigree. We have a lot of work ahead of us. Join us.

About the role

We’re hiring a Research Engineer to push Console’s core agent loop from powerful to truly self-improving. As more customers rely on Console to automate critical back-office operations, we need engineers who can turn real production traffic into better agents, evals, and specialist models.

This role sits at the intersection of applied AI engineering and research. You’ll build the systems that let us measure, debug, and improve agent behavior in real-world conditions: production traces, offline replay, labeling workflows, eval harnesses, prompt and program optimization, and fine-tuning loops for high-volume agent tasks.

One of your first major focus areas will be improving how Console’s agents reason over complex enterprise context. Our agents need to understand users, apps, devices, tickets, licenses, policies, and customer-specific data well enough to answer questions and take action reliably. You’ll help turn this into a compounding optimization loop: using production traces, evals, assisted labeling, prompt/program optimization, and targeted model adaptation to make the system measurably better over time.

You’ll work closely with product, engineering, and leadership to ship improvements into production quickly, while helping define what research at Console looks like as we scale.

What You’ll Do

You’ll build and improve AI systems that operate in real-world conditions, where reliability, speed, cost, and adaptability matter. At Console, you will:

  • Build the eval and optimization loop for Console’s core agents, turning real production usage into measurable improvements in quality, latency, and reliability

  • Systematically improve agent behavior across prompts, programs, routing logic, constraints, and model adaptations, applying techniques like DSPy and GEPA where useful

  • Fine-tune, adapt, and evaluate specialist models for repeatable, high-volume agent tasks where Console has clear production feedback or verifiable quality signals

  • Work across the stack when needed, from traces and eval infrastructure to agent orchestration, product workflows, and customer-facing AI behavior

Who You Are

You’re a product-minded research engineer who enjoys building real systems people depend on. You’ll likely have:

  • Strong technical background in software engineering, machine learning, or applied AI, demonstrated through an advanced degree and/or equivalent experience building production AI systems

  • Strong software engineering fundamentals and good judgment for designing, building, and debugging complex systems

  • Experience building evals for AI systems, including datasets, judges, metrics, offline replay, tracing, or regression testing

  • Practical understanding of modern model adaptation and post-training methods, including LoRA/QLoRA, SFT, distillation, preference optimization, reward modeling, DPO/GRPO, and reinforcement learning from verifiable feedback

  • Ownership mindset: you drive projects end-to-end, move quickly from real usage, and care about shipping measurable improvements

  • Enjoy following SOTA research into new models, agent architectures, evals, post-training methods, and optimization techniques

Why join Console?
  • Product-market fit: We have built the leading product in our category, in a massive market. We've hit an inflection point and are on track to build a generational company.

  • World-class team: We seek high agency contributors who are comfortable navigating ambiguity, ruthlessly prioritize what matters and are action-biased.

  • Grow with us: We reward impact, not credentials or years of experience. We intend to grow talent from within as we scale up.

  • Competitive pay and benefits: top compensation with full benefits including:

    • Comprehensive health, dental, and vision insurance

    • Unlimited PTO

    • 401(k)

    • Equity

    • Meals provided daily in office

HQ

Console San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

7 Days Ago
Hybrid
128K-160K Annually
Mid level
128K-160K Annually
Mid level
Artificial Intelligence • Hardware • Software • Nanotechnology • Semiconductor • Quantum Computing • Defense
Lead research and development of agentic AI and multi-agent systems integrating LLMs, memory, planning, tool use, and GraphRAG-style retrieval. Develop graph machine learning and geometric deep learning pipelines, build knowledge-augmented AI (knowledge graphs, ontologies), and ensure trustworthy AI via XAI, V&V, robustness testing, and uncertainty quantification. Collaborate across teams, publish research, and support proposal development for mission-critical autonomous and decision-support applications.
Top Skills: Agent2Agent (A2A)AgentopsAutogenCypherDistributed InferenceGeometric Deep LearningGnnsGpu AccelerationGraphragKnowledge GraphsLanggraphLlmopsLpgModel Context Protocol (Mcp)Neo4JOntologiesPythonPyTorchRayRdfSglangSparkVllm
8 Days Ago
Remote or Hybrid
35 Locations
100K-145K Annually
Mid level
100K-145K Annually
Mid level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Design, build, and maintain distributed cloud services to support advanced threat detection. Collaborate with security, data science, and engineering teams to integrate capabilities into the Falcon platform, ensure production reliability, and leverage AI/LLMs to improve detection and automation. Mentor peers and contribute to scalable, high-availability system architecture.
Top Skills: AWSAzureBambooCassandraDockerElasticsearchFalcon PlatformFlinkGCPGoGrafanaJavaJenkinsKafkaKubernetesLlmsPythonRedisRestful ApisScala
20 Days Ago
In-Office
Sunnyvale, CA, USA
207K-275K Annually
Senior level
207K-275K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead applied research to advance continuous learning for agents: design and evaluate LLM post‑training and RL methods, implement and deploy experiments at scale, validate research on customer tasks, optimize GPU/distributed training, and mentor engineers while driving cross‑functional technical direction.
Top Skills: CudaFastapiGpuKubernetesMegatronPostgresTemporal

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account