Description
Recognized as the No. 1 site trusted by real estate professionals, Realtor.com® has been at the forefront of online real estate for over 25 years, connecting buyers, sellers, and renters with trusted insights and expert guidance to find their perfect home. Through its robust suite of tools, Realtor.com® not only makes a significant impact on the real estate industry at large, but for consumers, navigating the biggest purchase they will make in their life, by providing a user experience that is easy to use, easy to understand, and most of all, easy to make decisions.
Join us on our mission to empower more people to find their way home by breaking barriers to entry, making the right connections, and building confidence through expert guidance.
We are seeking an AI Engineer to join our AI Integrations Team with deep expertise in AI safety, guardrails, and LLM evaluation to join our AI Integrations team. You will be responsible for building and tuning the classifiers that keep our AI assistant compliant, safe, and helpful - working at the intersection of AI and consumer-facing products.
About the Role:
The AI Integrations Team is a cross-functional squad that builds the future of how Realtor.com integrates AI across its products. Our flagship is RealAssist, an LLM-powered real-estate assistant embedded across web and mobile. Because it advises consumers on housing, it operates under real Fair Housing Act compliance obligations - our AI safety layer is a launch gate, not an afterthought. This role owns that layer, including:
- LLM-as-a-judge classifiers for Fair-Housing compliance and content moderation The guardrail evaluation harness that proves it all works before it ships
- A cloud prompt-injection / jailbreak backstop (Google Cloud Model Armor) Technical Ownership
- Build, tune, and own LLM-as-a-judge classifiers for Fair-Housing compliance and content moderation, targeting high recall on disallowed content without overblocking legitimate users
- Design and run guardrail evaluation pipelines: curate labeled datasets, maintain train/test/validate splits, and run offline prod-replay overblocking evals -
- Report confusion matrices and precision/recall/F1 per category to make safety decisions defensible
- Integrate and operate cloud guardrail services as a runtime prompt-injection/jailbreak screening layer, with fail-open behavior and alerting
- Red-team the assistant and wire safety-regression checks into CI so guardrails don't silently degrade as the product grows
- Manage guardrail infrastructure as code with Terraform (templates, IAM, project shape) - Participate in technical design reviews and architecture discussions
- Partner with ML, backend, product, and legal/compliance stakeholders to define risk tiering and get new AI capabilities through safety review before launch
What You'll Bring:
- AI Safety & Evaluation Expertise (Required)
- 4+ years of professional software/ML engineering experience, with hands-on LLM application work
- Bachelor’s degree or equivalent experience
- Strong Python proficiency
- Real classification-metrics literacy - precision/recall/F1 (macro vs weighted), confusion matrix debugging, dataset curation, and calibration to a production distribution -
- Experience building or tuning LLM prompts/classifiers and evaluating them with an eval framework (DeepEval, RAGAS, Arize Phoenix, LangSmith, OpenAI Evals, or a home grown harness)
- Experience turning small seed sets into robust labeled eval datasets (dataset-synthesis tooling, prompt-optimization loops)
- Familiarity with prompt-injection / jailbreak defense concepts (OWASP LLM Top 10, input/output filtering, least-privilege tool access, adversarial testing)
- Guardrails & Infrastructure Experience (Bonus)
- Hands-on with a cloud guardrail service (Google Cloud Model Armor, AWS Bedrock Guardrails, Azure AI Content Safety)
- Terraform / IaC for cloud infrastructure
- Experience with Google Cloud Vertex AI / Gemini
- Understanding of monitoring and observability tools (New Relic or similar)
- Exposure to regulated / compliance-sensitive domains (fair housing, fair lending, healthcare, finance, trust & safety)
- Red-teaming or AI-security background
Expectations
Technical Excellence
- Independently design, implement, and tune guardrail and evaluation systems - Make sound architectural decisions that balance safety, latency, and user experience - Write clean, maintainable, well-tested code that sets the standard for the team - Keep eval datasets calibrated to real traffic and emerging attack patterns - Stay current with AI-safety research, jailbreak techniques, and new guardrail tooling
Impact & Ownership
- Own the AI safety layer end-to-end from design through deployment
- Anticipate problems and proactively address them through red-teaming and regression checks
- Make data-driven decisions grounded in eval metrics
- Deliver high-quality work consistently on schedule
Why Join Us:
Cutting-Edge Technology
- Work with the latest AI technologies from leading providers
- Build the safety layer for LLMs running in production
- Shape the future of safe, compliant AI-powered real estate search at scale - Millions of users rely on Realtor.com to find their next home
- Your work directly affects major life decisions for consumers
- Keep the AI that guides those decisions fair, safe, and trustworthy Growth
- Work alongside Staff and Principal engineers who will challenge and support you - Exposure to full-stack technologies including mobile, backend, and AI - Clear path to career growth with opportunities for technical leadership
How We Work:
We balance creativity and innovation on a foundation of in-person collaboration. For most roles, our employees work four or more days in our offices, where they have the opportunity to collaborate in-person, adding richness to our culture and knitting us closer together.
How We Reward You:
Realtor.com® is committed to investing in the health and wellbeing of our employees and their families. Our benefits programs include, but are not limited to:
- Inclusive and Competitive medical, Rx, dental, and vision coverage
- Family forming benefits
- 13 Paid Holidays
- Flexible Time Off
- 8 hours of paid Volunteer Time off
- Immediate eligibility into Company 401(k) plan with 3.5% company match
- Tuition Reimbursement program for degreed and non-degreed programs
- 1:1 personalized Financial Planning Sessions
- Student Debt Retirement Savings Match program
- Free snacks and refreshments in each office location
Do the best work of your life at Realtor.com®
Here, you’ll partner with a diverse team of experts as you use leading-edge tech to empower everyone to meet a crucial goal: finding their way home. And you’ll find your way home too. At Realtor.com®, you’ll bring your full self to work as you innovate with speed, serve our consumers, and champion your teammates. In return, we’ll provide you with a warm, welcoming, and inclusive culture; intellectual challenges; and the development opportunities you need to grow.
Diversity is important to us, therefore, Realtor.com® is an Equal Opportunity Employer regardless of age, color, national origin, race, religion, creed, gender, sex, sexual orientation, gender identity and/or expression, marital status, status as a disabled veteran and/or veteran of the Vietnam Era or any other characteristic protected by federal, state or local law. In addition, Realtor.com® will provide reasonable accommodations for otherwise qualified disabled individuals.
Similar Jobs at Realtor.com
What you need to know about the San Francisco Tech Scene
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

