logoalt Hacker News

oerstedtoday at 6:50 AM10 repliesview on HN

There's this dilemma where in theory there's a ton of demand for engineers that can do real LLM machine-learning, but in practice there are very few available positions and entrepreneurship opportunities.

The reality is that an incredibly small minority of companies in the world do any real training or optimisation. It's unnecessary and inefficient for most purposes unless you are fully dedicated to being an LLM company, and still then it's a struggle. Those few that do train, they spend most of their budget on compute and have relatively small teams.

Getting experience in this field requires having access to very expensive hardware to begin with. And the skills will be quite hard to convert into any real value for someone, leading to a decent income, unless you have a ton of funding from patient investors, or you have decent contacts in Bay Area networks to get hired at the right place.

With all due respect, paulg is in somewhat of a bubble, this is not congruent with the global situation.


Replies

yobbotoday at 8:41 AM

Yes. It's like looking at the (Apollo) moon rocket launch and then suggesting teenagers should learn to build rockets in their garages for the coming space age.

It is viable as a toy project, but there are vanishingly few career opportunities.

show 2 replies
kevmo314today at 6:52 AM

That's like saying the only way to do real engineering is with Google-scale Borg deployments. You can do quite a lot on very little hardware, r/StableDiffusion is a prime example.

show 3 replies
khalictoday at 9:44 AM

lol unnecessary and inefficient…

I just finished fine tuning Gemma e2b for local code completion on my local machine.

This comment just reinforces what the post actually means. We need people that are LLM natives, computing solves itself with time and with scale adjustments

show 1 reply
Zyloklototoday at 8:03 AM

We finetune LLMs. Small ones like Gemma 4 for semantic tasks.

There are plenty of areas were we need people to do this for insurances, banks etc.

AI/ML exists on many levels.

show 1 reply
epolanskitoday at 8:10 AM

+1, I have a friend, math PhD that's been working on ML research 5+ years in London yet he has not been able to find any position.

The only jobs that he found he was highly over qualified or paid very little.

In any case, it doesn't look like there's this crazy rush to hire all ML talent, even the one that understand the math and technology deeply.

show 3 replies
willtemperleytoday at 7:18 AM

I think companies of all sizes will want their own models, or at least customised ones, for their own specific use cases or competition and security issues.

1. Both training and optimisation will get significantly cheaper and easier quickly.

2. Politics will probably get even more insane before a potential reprieve on the 20th of Jan 2029.

3. The big AI firms will become part of the surveillance capitalism network, if they're not already.

So I think for self-protection a lot of companies will be looking near to medium term AI independence.

show 4 replies
joshuakcockrelltoday at 8:26 AM

This is like saying, “teens shouldn’t learn how to make their own game engine because no one is hiring for that.”

You’re missing the point. Understanding how Unity works fundamentally makes you a better Unity dev.

show 2 replies
dukeyukeytoday at 7:42 AM

Did you read his comments on this? It's not to actually do LLM research stuff, it's to trigger and unlock ideas.

spwa4today at 6:57 AM

That's not true, because everyone, everyone, everyone seems to want to do training. Which results in a 50 person company training, say, a voice model that then fails, because it's just not good enough.

In reality the problem is that it gets blasted out of the water by a much worse architecture trained on 10000x the infrastructure. And while I'm sure the freshly brought in ML student came up with a 10%, even 30% better architecture, it just doesn't matter. (and never mind that even OpenAI hasn't really solved a voice model yet. Try it. It can probably match 2026-quality call centers, but it's no substitute for an actually empowered human)

... and yet, if you look at what hyperscalers are getting paid for ... comfortably more than half the income is training. Which makes no sense on so many levels.

e.g. https://valueaddvc.com/blog/inference-chips-vs-training-chip... (I get it, not great first source, but st

show 2 replies
aaron695today at 8:18 AM

[dead]