Hire Machine Learning developers in San Diego
What the published figures say about hiring this skill into the San Diego-Chula Vista-Carlsbad, CA market, what a seat here actually costs, and how to run the search so it closes.
The local market for this occupation
Machine Learning developers are counted by the Bureau of Labor Statistics under Data Scientists. In the San Diego-Chula Vista-Carlsbad, CA area the Bureau puts employment in that occupation at 2,830, with a median annual wage of $130,990. The national median for the same occupation is $120,230, so San Diego sits about 9% above the country as a whole.
Ranked against every US metropolitan area where this occupation is separately published, San Diego is 19 of 282 by median wage. Of the 28 markets covered on this site it is 9. Those two numbers together are a better guide to what an offer needs to look like than any single national figure, because they say where the market sits rather than what it averages.
The middle half of the local market runs from $93,420 to $164,160. That band, not the midpoint, is the number to carry into a budget conversation about a machine learning engineer here. Where in the band a particular hire lands depends far more on what the person can be left to own than on how many years the CV shows, which is why the level definitions matter more than the title.
One qualification applies to all of this and it is worth stating plainly. The Bureau classifies by occupation, not by technology. Nobody publishes an official wage figure for Machine Learning specifically, and any site that quotes one has either modelled it or made it up. The occupation figures are the honest available baseline: they describe the market a machine learning engineer is hired into, and the technology adjusts where inside that band a given person sits.
What each level looks like against local pay
Percentiles describe a market; they do not describe a person. The useful step is to read the published Data Scientists band for San Diego against what someone at each level can actually be left to own. The mapping below is a working guide rather than a rule, and the overlap between adjacent levels is genuine: a strong mid-level engineer can be worth more than a weak senior one and frequently is.
| Level | Indicative local band | What they can be left to own |
|---|---|---|
| Junior | $74,110 to $93,420 | Trains and evaluates models on prepared data. Needs review on evaluation design and leakage. |
| Mid-level | $93,420 to $130,990 | Owns a model end to end including its pipeline, deployment and monitoring. Designs evaluation that reflects the business cost. |
| Senior | $130,990 to $164,160 | Owns system architecture, the feature and training infrastructure, the monitoring and retraining strategy, and can say when machine learning is not the answer. |
| Staff | $164,160 to $201,190 | Owns the platform across teams, the standards for evaluation and deployment, and the judgement about which problems merit models at all. |
Bands are percentiles of the published local occupation wage, not FuturByte rates, and not a guarantee that any individual sits where the band suggests.
Two things go wrong when this mapping is used carelessly. The first is budgeting a senior seat at the local median and then interviewing people who can genuinely own an area of the system; in San Diego that band starts around $130,990 and an offer below it will not close against a counter. The second is the reverse: paying at the top of the band for someone who still needs their work scoped by somebody else. The level definitions are there so the conversation is about what the person will own rather than about how many years the CV shows.
It is also worth being explicit that years and level are only loosely related in Machine Learning work. Somebody who has spent six years maintaining one application carries less transferable judgement than somebody who has spent three across three very different ones. When a CV and a band disagree, the interview should settle it, not the CV.
The titles you are bidding against locally
A machine learning engineer in San Diego is not only being recruited by other teams hiring the same title. The same person is a credible candidate for several adjacent occupations, and what those pay locally is part of what any offer has to clear. These are the published local figures for the titles that compete for this pool.
| Occupation | Employed | 25th percentile | Median | 75th percentile | 90th percentile |
|---|---|---|---|---|---|
| Data Scientists | 2,830 | $93,420 | $130,990 | $164,160 | $201,190 |
| Software Developers | 20,610 | $128,590 | $163,270 | $193,980 | $223,260 |
| Computer and Information Research Scientists | 1,680 | $96,180 | $121,530 | $156,480 | $195,190 |
| Web Developers | 520 | $79,250 | $98,300 | $135,570 | $174,430 |
| Software QA Analysts and Testers | 2,280 | $87,000 | $120,210 | $142,230 | $169,350 |
| Information Security Analysts | 1,330 | $102,320 | $132,120 | $167,480 | $206,510 |
| Computer and Information Systems Managers | 7,250 | $165,460 | $206,740 | $262,210 | $304,530 |
Where a row is missing, the Bureau does not publish a separate estimate for this metro, usually because the local sample is too small to release.
The spread across these titles in San Diego runs from $206,740 for computer and information systems managers down to $98,300 for web developers. That gap is the practical reason technical people move sideways between titles rather than up within one: in this market the fastest available pay rise for a competent engineer is often a change of job title rather than a change of employer. If you are hiring at the lower end of that range, expect to lose some candidates to the upper end of it, and expect that to happen after they have accepted.
This also affects how a role should be written. A specification that describes the work in terms of one narrow title competes only for people who already hold it. One that describes the system and the ownership on offer reaches people currently sitting under a different title who would be entirely capable of the work, and that is usually where the available capacity in a tight market actually is.
Which direction this market is moving
A single year tells you the price. Three tell you whether it is going up. These are the published figures for data scientists in the San Diego-Chula Vista-Carlsbad, CA area across the last three releases.
| Reference period | Employed | Median wage |
|---|---|---|
| May 2023 | 2,340 | $127,300 |
| May 2024 | 2,550 | $127,200 |
| May 2025 | 2,830 | $130,990 |
Source: BLS Occupational Employment and Wage Statistics, metropolitan area files. The Bureau does not design these releases to be read as a time series; treat the movement as a direction rather than as a growth rate.
Across those two years the local median moved up 3%. Employment moved up 21% over the same period. For someone planning a machine learning engineer hire in San Diego, the practical reading is that a salary band set from figures more than a year old is probably still close enough to work from.
Markets a candidate here would also consider
Candidates do not compare your offer against a national average. They compare it against what they could get nearby, and for a machine learning engineer in San Diego that means a handful of specific metros. These are the closest comparisons on the published figures.
| Metro area | Employed | Median wage | vs San Diego | Location quotient |
|---|---|---|---|---|
| San Diego-Chula Vista-Carlsbad, CA | 2,830 | $130,990 | — | 1.09 |
| Los Angeles-Long Beach-Anaheim, CA | 9,850 | $129,740 | -1% | 0.93 |
| San Jose-Sunnyvale-Santa Clara, CA | 6,060 | $185,080 | +41% | 3.16 |
| San Francisco-Oakland-Fremont, CA | 10,460 | $170,110 | +30% | 2.61 |
Percentages compare each metro against San Diego rather than against the national median.
San Jose-Sunnyvale-Santa Clara and San Francisco-Oakland-Fremont pay meaningfully more for this occupation than San Diego does. For remote-capable work that is a real competitor for the same people, and it is worth knowing before you set a band rather than after a candidate declines.
The same table read the other way is a sourcing map. If the local pool is thin and a neighbouring metro pays less for the same occupation, that metro is where a remote or relocating candidate is most likely to come from, and the conversation is easier because the move is upward for them.
Where Machine Learning work actually sits in this economy
The San Diego economy is anchored by wireless and telecom, biotech, defense and genomics. Machine Learning is not used identically across those, and the version of the skill that is abundant locally is shaped by whichever of them employs the most engineers. That is the part a national salary table cannot tell you and it is usually what decides whether a shortlist converts.
Wireless and telecom
Where Machine Learning appears in wireless and telecom, it most often looks like the computer vision pattern: Inspection, medical imaging and document processing, where data labelling is usually the dominant cost. Developers coming out of this part of the San Diego market therefore tend to arrive strong on the constraints that sector imposes and lighter on the ones it never had to deal with. If your product shares those constraints, that is experience you would otherwise spend a year building. If it does not, the gap is real, and it is a fair thing to ask about directly rather than to discover in month two.
Biotech
Where Machine Learning appears in biotech, it most often looks like the demand and forecasting pattern: Inventory, pricing and capacity planning, where time-aware evaluation is essential and often mishandled. Developers coming out of this part of the San Diego market therefore tend to arrive strong on the constraints that sector imposes and lighter on the ones it never had to deal with. If your product shares those constraints, that is experience you would otherwise spend a year building. If it does not, the gap is real, and it is a fair thing to ask about directly rather than to discover in month two.
Defense
Where Machine Learning appears in defense, it most often looks like the demand and forecasting pattern: Inventory, pricing and capacity planning, where time-aware evaluation is essential and often mishandled. Developers coming out of this part of the San Diego market therefore tend to arrive strong on the constraints that sector imposes and lighter on the ones it never had to deal with. If your product shares those constraints, that is experience you would otherwise spend a year building. If it does not, the gap is real, and it is a fair thing to ask about directly rather than to discover in month two.
Genomics
Where Machine Learning appears in genomics, it most often looks like the computer vision pattern: Inspection, medical imaging and document processing, where data labelling is usually the dominant cost. Developers coming out of this part of the San Diego market therefore tend to arrive strong on the constraints that sector imposes and lighter on the ones it never had to deal with. If your product shares those constraints, that is experience you would otherwise spend a year building. If it does not, the gap is real, and it is a fair thing to ask about directly rather than to discover in month two.
The practical use of this is in reading CVs rather than in sourcing. Two candidates in San Diego with the same number of years of Machine Learning can have been solving quite different problems, and the interview should be aimed at the difference rather than at the technology they have in common.
What a local machine learning engineer seat costs to keep open
Salary is the quoted number and it is not the budget. Below is the employer-side payroll cost of one person in this occupation in San Diego, using published local wages and the statutory 2026 employer rates.
| Wage point | Annual wage | Employer OASDI and Medicare | Wage plus these taxes |
|---|---|---|---|
| 25th percentile | $93,420 | $7,147 | $100,567 |
| Median | $130,990 | $10,021 | $141,011 |
| 75th percentile | $164,160 | $12,558 | $176,718 |
Employer OASDI at 6.2% to a wage base of $184,500, Medicare at 1.45% uncapped. Source: Social Security Administration, Contribution and Benefit Base, retrieved 2026-09-25. Unemployment insurance, benefits, equipment and recruitment cost are additional.
Then there is the cost nobody puts in the model. Every month the role is open, the work it was meant to do is not happening. That cost is the same whichever way you eventually fill the seat, and in a specialist search it routinely exceeds the difference between the options being compared. It is the single most common reason a cost comparison that looked careful turns out to have been wrong.
What to test when the pool is this one
The full interview guide for this technology is on the Machine Learning developers page. What changes in San Diego is emphasis rather than substance: given what the local market has been building, these are the areas where candidates here differ most from each other, and therefore where an interview earns its keep.
A model that failed in production
Operational judgement in this field comes from failures.
- Strong answer: Describes what degraded, how it was detected, and what changed afterwards.
- Warning sign: Has never had a model in production.
How they evaluate a model
The single most informative question, and where weak candidates are exposed immediately.
- Strong answer: Chooses metrics that match the business cost of each error type, splits appropriately including by time, and holds out a genuine test set.
- Warning sign: Quotes accuracy on an imbalanced problem, or tunes against the test set.
Data leakage
The most common reason a model performs well in testing and badly in production.
- Strong answer: Can describe a leak they found, and checks for it systematically rather than hoping.
- Warning sign: Unfamiliar with the concept, or has never had a model underperform after deployment.
Two patterns worth asking about directly, because they show up in inherited codebases far more often than candidates volunteer them:
Training and serving skew
- What you see: Feature computation implemented separately in the training pipeline and the serving path.
- What it costs: Predictions quietly worse than evaluation suggested, with no error to alert anyone.
- The fix: Share the transformation code, or use a feature store. Verify with the same input through both paths.
Notebooks in production
- What you see: Training or scoring code run from a notebook on a schedule.
- What it costs: Hidden state, no tests, no reproducibility, and failures nobody else can debug.
- The fix: Move to tested modules with orchestration. Keep notebooks for exploration.
The options, and what actually decides between them
Once the local figures are on the table, most teams are choosing between three ways of getting Machine Learning capacity into San Diego. They are not ranked. Which one is right depends on how long the work lasts, how much of it there is, and how much of the surrounding context the person needs to hold.
Hiring locally onto your own payroll
The right answer when the work is permanent, when the person needs to accumulate context that has no value anywhere else, or when presence in San Diego is a genuine requirement rather than a preference. The costs are the ones in the table above plus benefits and recruitment, and the risk is time: at this concentration the pool is thin enough that a specialist search can run for months without producing a viable shortlist.
Adding a vetted developer to your existing team
Staff augmentation suits work that is real but not permanent, and teams that already have the review capacity and the architectural direction in place. The person joins your standups, your repository and your process. What you are buying is capacity and specific Machine Learning experience, not decision-making, and the constraint is almost always how much code your existing team can review rather than how many developers you add.
A dedicated team that owns an area
The right shape when there is a whole area of work to own rather than a queue of tickets, and when you would otherwise be hiring three or four people at once into a market where that takes a year. It asks more of you at the start, because an area cannot be owned without a clear definition of what it includes and who decides, and it asks less of you afterwards.
The comparison people get wrong is between a local salary and an hourly rate. Those are not the same quantity. A fair comparison puts the fully loaded employer cost of a local seat, including the months it stands empty and the recruitment spend that filled it, against the total cost of the alternative including the coordination overhead it adds. Run honestly, that comparison sometimes favours hiring locally, and when it does we will say so.
A realistic plan for this search
Decide early what genuinely has to be local
At this concentration the pool is the constraint. Separate the parts of the work that require presence in San Diego from the parts that do not, and run those as two different searches. Most teams discover the genuinely local list is shorter than they assumed.
Widen before you wait
Extending a local-only search by a quarter costs a quarter of output. Widening the geography costs a conversation about how the team works. The second is almost always the cheaper trade.
Write the brief around the system, not the stack
State what the system does, what it runs on, what is already decided and what the new person would own. A list of technologies with no context filters for keyword matches, and keyword matches are exactly who gets rejected in the technical round.
Use the local band, not a national median
The published middle half for this occupation in San Diego is the range offers actually land in. Anchoring on a national figure produces an offer that is either uncompetitive or unnecessarily expensive, and you will not always find out which.
What changes when the developer is not in the building
Most of what makes a distributed machine learning engineer productive is decided in their first two weeks, and almost all of it is on the client side. These are the things that reliably separate a first merged change in week one from a first merged change in week four.
Decisions written down where they can be found
In a team split across time zones, a decision made in a conversation in San Diego does not exist for anyone who was not in it. This is the discipline that distributed teams either build early or pay for repeatedly.
An agreed overlap window
A few hours, published, treated as real, and used for review and decisions rather than status. Teams that skip this do not save meeting time; they spend it several times over in waiting.
A running environment on day one
Not documentation describing how to build one. An environment that starts, with seed data, on the machine the developer actually has. Every day spent on environment setup is a day billed at full rate for no output, and it is the single most common avoidable cost in an engagement.
A named reviewer with real capacity
Someone whose job explicitly includes reviewing this work, not someone who will get to it. A developer who waits two days for review does a quarter of the work they otherwise would, and the cost of that lands on you.
None of this is specific to working with us. It is what any developer joining any team needs, and it is worth stating plainly because the failures above get attributed to the developer far more often than to the setup that produced them.
The same role in other US markets
Ranked by how many people are employed in this occupation locally, which is the figure that most affects how long a search takes.
- Machine Learning developers in New York $135,980
- Machine Learning developers in San Francisco $170,110
- Machine Learning developers in Dallas-Fort Worth $127,750
- Machine Learning developers in Los Angeles $129,740
- Machine Learning developers in Washington, D.C. $132,200
- Machine Learning developers in Seattle $164,740
- Machine Learning developers in Chicago $107,640
- Machine Learning developers in Boston $132,040
- Machine Learning developers in Atlanta $108,940
- Machine Learning developers in Philadelphia $109,910
- Machine Learning developers in San Jose $185,080
- Machine Learning developers in Denver $112,520
- Machine Learning developers in Charlotte $132,460
- Machine Learning developers in Houston $106,750
- Machine Learning developers in Detroit $103,330
Other technologies in San Diego
Skills that appear alongside Machine Learning on most job specifications.
- AI Engineering developers in San Diego
- Python developers in San Diego
- Data Engineering developers in San Diego
- Data Science developers in San Diego
- AWS developers in San Diego
For the technology itself, including the full interview guide, migration paths and what each level can own, see hiring Machine Learning developers. For the wider San Diego technical market across every occupation, see hiring developers in San Diego.
Frequently asked questions
What does a machine learning engineer cost in San Diego?
There is no official wage figure for Machine Learning specifically, because the Bureau of Labor Statistics classifies by occupation rather than by technology. The honest baseline is Data Scientists in the San Diego-Chula Vista-Carlsbad, CA area, where the median annual wage is $130,990 and the middle half of the market runs from $93,420 to $164,160. Those are employer wages, before payroll taxes, benefits and recruitment cost.
How many machine learning engineers are there in San Diego?
Nobody counts developers by technology, so any specific number you see quoted is an estimate. What is published is the occupation: 2,830 people in data scientists in the San Diego-Chula Vista-Carlsbad, CA area. Machine Learning is one technology inside that population, and the share using it is a matter of inference rather than record.
Is it faster to hire a machine learning engineer locally in San Diego or remotely?
At this concentration the local pool is the constraint rather than the competition, so a local-only search for a specific technology tends to run long. Widening the geography usually shortens the calendar more than any change to the process will.
Do we need someone in the San Diego time zone?
Usually less than teams assume. What genuinely needs the local clock is live incident response, work with people who are only available in local hours, and anything tied to a physical site. Everything else needs a committed overlap window of a few hours rather than a matching working day. San Diego runs on Pacific time.
Which local industries will a machine learning engineer here have come from?
The anchors of this economy are wireless and telecom, biotech, defense and genomics. Most experienced candidates in this market will have spent time in at least one of them, and that background shapes both what they are good at and what they have never had to handle. It is worth asking about explicitly rather than inferring from the CV.
Can you supply a machine learning engineer who overlaps with San Diego hours?
Yes. Overlap is the thing we schedule around rather than a side effect of where someone happens to live, and it is agreed before an engagement starts rather than negotiated afterwards. Tell us which hours genuinely need to be covered and why, and we will tell you whether we can meet it.
How do you assess a machine learning engineer before we see them?
Working code and a conversation about decisions, not a quiz. We look at what someone has built, ask what they would now do differently and why, and probe the areas where this technology most reliably separates people. You see our reasoning alongside the shortlist, including the reservations.
What if the shortlist is wrong?
Tell us why and we will recalibrate. A rejected shortlist normally means the brief and the need had drifted apart, and that is worth finding out in week one rather than month three. We would rather say we are not the right fit for a role than keep sending candidates against a brief that is not working.