
Anthropic · San Francisco, CA
ABOUT ANTHROPIC Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for ...
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our
users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and
business leaders working together to build beneficial AI systems.
Anthropic is at the forefront of AI research, dedicated to developing safe, ethical, and powerful artificial intelligence. Our
mission is to ensure that transformative AI systems are aligned with human interests. We're looking for an experienced tech lead
to join our Evals Infrastructure team, building the systems that let us measure what our models can actually do. Evaluation is how
we know whether a model is safe to ship — you'd own the infrastructure that makes those measurements fast, reliable, and
trustworthy at scale. In this role you'll work at the intersection of inference, research and infrastructure engineering: managing
the large scale distributed systems that orchestrate evals for our frontier models, building and scaling the harnesses researchers
use to design and run evals, making results reproducible and interpretable, and ensuring eval signal is available where decisions
get made. Your work directly shapes what we build and what we don't.
Responsibilities
reuse of eval work
You may be a good fit if you
tech-lead-with-reports experience)
Strong candidates may have
Sample Projects
decision-grade
The annual compensation range for this role is listed below.
For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales
commissions/sales bonuses target and annual base salary for the role.
Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some
roles may require more time in our offices.
Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate.
But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help
with this.
We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet
every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more
prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself
prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have
enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range
of diverse perspectives on our team.
Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you
from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as
working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for
money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any
links—visit anthropic.com/careers [http://anthropic.com/careers] directly for confirmed position openings.
We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few
large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work
on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and
biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research
discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication
skills.
The easiest way to understand our research directions is to read our recent research. This research continues many of the
directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling
Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional
equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to
collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy
[https://www.anthropic.com/candidate-ai-guidance] for using AI in our application process.
ABOUT ANTHROPIC Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. ABOUT THE ROLE Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We believe powerful AI can compress a century of scientific progress into a decade. Claude Science is a big part of building toward that future. Claude Science [https://www.anthropic.com/news/claude-science-ai-workbench] is an AI workbench that gives researchers a single environment for work that today spans dozens of disconnected tools — literature, specialized databases, scientific computing, analysis, and publication-ready outputs. Claude Science already renders protein structures, genome tracks, and chemical structures natively, and coordinates multi-agent workflows with built-in review for citation and calculation errors — and we’re just getting started. Much of this is being built 0→1 right now: you’ll shape both the product and the architecture in a category no one has defined yet. You'll be a technical leader who thinks holistically about the end-to-end researcher experience, partners directly with our internal research team to push model capabilities into production, and carries real ownership over what we ship next. Our north star is accelerating scientific progress — across biology, chemistry, physics, and beyond — from early discovery through real-world application, by an order of magnitude. This work supports programs like Anthropic’s AI for Science Program [https://www.anthropic.com/news/ai-for-science-program] and our rare disease research grants [https://www.anthropic.com/news/rare-disease-research-grants]. WHAT YOU'LL DO * Ship fast against a roadmap you help shape: this is a product in a category no one has defined yet, and the highest-leverage problems are still unclaimed * Interface directly with working scientists — academic labs, industry R&D teams, and research institutes — during key conversations, translating what you learn into engineering priorities * Partner with product and design to turn how scientists actually work — from hypothesis to analysis to publication — into shipped product * Work closely with research to make the models better at science: shaping evals, surfacing failure modes, and feeding what users hit in the real world back into model development YOU MAY BE A GOOD FIT IF YOU * Have 8+ years of software engineering experience, ideally with 2+ years at a Staff or equivalent technical leadership level * Have built products from 0 to 1 in fast-moving environments, and can set technical direction with limited precedent to lean on * Have built AI products and know what it takes to turn model capabilities into applications people actually use * Are comfortable working directly with technical domain experts and translating what you learn * Drive cross-team alignment to ship impactful work, with influence over authority STRONG CANDIDATES MAY ALSO HAVE * Background in chemistry, biology, physics, or another science * Experience working with research teams to improve domain-specific model capabilities, including evaluation frameworks The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $405,000—$485,000 USD LOGISTICS Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers [http://anthropic.com/careers] directly for confirmed position openings. HOW WE'RE DIFFERENT We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. COME WORK WITH US! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy [https://www.anthropic.com/candidate-ai-guidance] for using AI in our application process.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. ABOUT THE ROLE The Cortex Apps team is building the future of AI for enterprise data. See our flagship product in action: Talk To Your Data: Snowflake Intelligence Demo. Your work will directly impact how businesses understand and leverage their data. You'll own the full AI engineering lifecycle: design, prompt/tool engineering, evals, deployment, measurement, and optimization. You'll work with a small, high-powered modeling and infrastructure team. WHAT YOU WILL DO IN THIS ROLE: * Own features end-to-end for Snowflake Cortex products. Build agentic workflows, NL-to-SQL on semantic layers, search. * Build enterprise-grade context engineering: function calling, tool schemas, guardrails, semantic model-aware prompting for SQL, and verification/repair * Design evals and hillclimb : create golden sets, create rubrics and metrics, analyze errors, run experiments to hill climb on the metrics. * Partner with product and infra: translate customer problems into products and experiments. Collaborate with infrastructure teams to productionize improvements. * Lead a team of engineers towards building great products REQUIREMENTS: * Bachelor’s degree in Computer Science, Engineering, Statistics or a related field. Master’s or higher degree preferred but not a requirement. * 8+ years of experience shipping AI features in production. * Proficiency in programming languages such as Python, Go * Strong communication skills and ability to collaborate effectively in a team environment. * (Optional) Experience working with text2sql, data modeling, data analysis, retrieval systems, and semantic layers is a plus. ABOUT SNOWFLAKE Snowflake is the AI Data Cloud trusted by the world's most innovative companies. We're shipping production-ready AI applications at scale and want you to join us in building the future of how businesses interact with their data through our Cortex products: Snowflake intelligence, cortex agents, cortex analyst, cortex search. Every Snowflake employee is expected to follow the company’s confidentiality and security standards for handling sensitive data. Snowflake employees must abide by the company’s data security plan as an essential part of their duties. It is every employee's duty to keep customer information secure and confidential. Snowflake is growing fast, and we’re scaling our team to help enable and accelerate our growth. We are looking for people who share our values, challenge ordinary thinking, and push the pace of innovation while building a future for themselves and Snowflake. How do you want to make your impact? For jobs located in the United States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.com
We believe the next era of how a company operates is one where humans and always-on agents work side by side - and we're building the platform that gets Melio there. We're looking for an Engineering Manager to lead AI Platform Enablement: a small, high-leverage team that brings agentic teammates to every function at Melio. You'll reshape how the org works day to day, leading senior engineers who build the runtime, frameworks, and patterns that let people across Finance, Operations, CX, Risk, Engineering, and beyond stand up always-on agents that work alongside them. Qualifications: * 4+ years managing engineering teams at scale, with a track record of shipping production AI or LLM-powered features * Deeply technical and passionate about the AI landscape - LLMs, agents, evals, prompting, the trade-offs between frameworks * Strong product instincts - you've worked closely with PMs, designers, and users, and you make calls about what to build based on outcomes, not output * Experience driving org-wide adoption of a new technology - cloud migration, observability rollout, AI tooling, anything where the technical work was the easy part * A bias toward adoption and measurable impact, energized by driving org-wide change in a space with no established playbook to inherit - you build the playbook * Comfortable operating across technical and non-technical audiences; you can hold your own with engineers on architecture trade-offs and with Finance or Ops leaders on workflow change Bonus points: * You've built or led a platform team whose primary customers were non-engineers * You've scaled LLM usage inside a company and have opinions about cost, governance, evals, and paved roads * You teach, write, or speak publicly about engineering, AI, or platform work A day in the life and how you'll make an impact: * Partnering with leaders across Finance, Operations, CX, Risk, and Engineering to identify the agentic teammates that will most change how their teams work, then prioritizing your roadmap around the ones that will land * Working with your tech lead and ICs to shape the platform itself: what capabilities it exposes, which patterns it standardizes, where it stays opinionated vs. flexible * Running adoption and education across Melio - documentation, workshops, office hours, success stories, and the unglamorous work of helping the second and third teams ramp up * Owning a small portfolio of custom agentic workflows your team delivers for high-leverage use cases, proving the platform's value with real, in-production agents * Measuring success in outcomes: teams actively using agentic teammates, work meaningfully shifted from humans to agents, and the rate at which new internal customers come online About the hiring department: The AI Platform Enablement team is Melio's bet that the next era of how a company operates is one where humans and always-on agents work side by side, and that getting there requires more than handing teams API keys. We build the platform that lets anyone at Melio - engineer or not - stand up an agentic teammate that does real work, persistently, alongside the humans it supports. We build custom agentic workflows for the use cases where the leverage is highest. And we own the harder half of the job: helping people across the company actually change how they work. We sit inside Backend Platform Engineering and partner closely with the Data & AI Infra team that owns model serving and infrastructure, with Data Apps on agentic data workflows, and with the company's AI Leads forum on direction. Our customers are everyone at Melio, which makes the work unusually broad and unusually impactful. About Melio: Melio, now part of Xero, is building the future of business payments for small businesses across America. Following our acquisition by Xero, the leading accounting software for small business, we're combining Melio's industry-leading bill pay solution with Xero's powerful accounting tools to transform how millions of small businesses manage their finances. As the fastest-growing B2B payment platform in the US, processing over $100B annually for more than 100,000 businesses, we're just getting started. Together with Xero, we're creating a seamless financial ecosystem that lets small business owners spend less time on back-office tasks and more time doing what they love. With offices in New York and Tel Aviv, Melio offers a vibrant, collaborative culture where your work directly impacts millions of small businesses. We're a diverse team of passionate individuals who believe in moving fast, thinking big, and building products that truly matter. If you're ready to help us dominate the US SMB market while being part of something transformative, we'd love to hear from you.