It's nice, I am just afraid of the pace Chris Lu can maintain with LLM assistance. Too many features are added in each new release, no stabilization time at all.
You haven't tried DeekSeek v4 or GLM 5.3 or Qwen 3.8 Next?
You are missing out a lot.
Try that with Hermes or Opencode or Deekseek Harness , even Qwen 3.8 27b works really well for that kind of that.
I just ask it to install windows as a vm on my linux and install vs Community 2019 on it , and then build a legacy vb 2019 project on it. and sleep
When i wake up :
It installs Qemu , setup a vm , inside vm download and install windows 10 on its own , clicking next next next as needed , typing in things , writing powershell , python scripts , that run automatically after install by baking into CD that includes ssh server , reboot , it logins into ssh , trigger pythons script that continue installation of vs 2019 community , which includes a driver that click the installation steps , installs nuget , install all depedencies and then build the project into exe after i woke up.
I've tried the latest Qwen, and without Internet, it still can go into an incoherent loop if you ask for, say, song lyrics. TBH the commercial models might do that too if not for their internal tooling.
Besides from privacy:
I already making twice now.you own the hardware and the price had doubled since i bought. Almost tripled.
You missed the opportunity and i have 4 of those awesome machines. Cry on.
I sell those to business who need local air gapped requirments and I make a lot more money!
I can run the alliterated models where none of the service prvoider even dare to provide.
THose benefits outweights a few K.
And show me an api provider that allows me to run 10x agents concurrently for 5 days straights .
> And show me an api provider that allows me to run 10x agents concurrently for 5 days straights .
Any of them on a Max/Pro plan as long as you are smart about model selection? That's my main objection to local inference, I'd need a whole rack of GPUs to do as many things in parallel that I can do for $400 a month. I do plan on setting up some local inference hardware, but...RAM and GPU prices alone are $$$$
I actually believe the amount of $ is what was missing for them to improve their fundamental flaws. They're a very competent team, but were working with a tiny fraction of the budget of US/China teams.
If Mistral aren't going to distill other people's large models they obviously need to train their own large models. This obviously requires money for optimization, tuning and training hardware.
They've started hosting GLM-5.2, that should be very telling of their capabilities at the moment. Hopefully the investment will allow them to hire the right people to become competitive.
Chinese models are trained on dubiously collected model traces from Claude/OpenAI that are purchased from model routers. All the major Chinese models use this data. That's one reason they've been able to catch up with Anthropic/OpenAI so quickly, they have so much data.
Mistral can't train on that data, because this data would be illegal to purchase & train on in the EU.
They have done that in the past. I clearly remember the times when they dropped releases as a simple tweet and it made the front page of HN - just like the recent Chinese models releases.
Those in front can and will change over time, and regions determine what is measured in the first place. So if they want, the EU can for example define "frontier" as "models made by a European nation that score highest on Math benchmarks".
There is logic to it. The ones who work 12h/day for 6d/week does more than the ones who work 8h/day for 5d/week does more than... It's simple scaling that pays off over time even if the general quality of work by the former isn't as high as that of the latter.
Shame them... for not working their employees into a heart attack at 40?
I'm not sure if your hypothesis is correct. I personally don't think the crazy work hours are what make the US so competitive in tech. But even if it were, it's odd to call for shaming countries for having less insane work cultures.
> We need to start shaming individual countries in addition to blanket EU.
Reading this gave me a visceral disgust reaction.
I have worked in the middle of the night here to be able to ship prod fixes for meeting deadlines. But that and even overtime should be exceptions rather than the norm.
This better be bait, in which case good job but flagged anyways. Surely nobody would genuinely make the argument that workers should be worked to the bone?
What I personally think could help would be less bureaucracy and regulation in regards to tech advancements (while preserving privacy of individuals), so that the time is used more efficiently, alongside significant investments.
Way to reduce a multi dimensional concept into a single scalar value. Care to explain what exactly make it inferior? Try using more than one word if you can.
Honestly being only 1 year behind makes me an optimist. You’re telling me Europe can be slightly behind with 1000x less capex and way more sustainable economics? Awesome. The world moves slower than AI progresses, I can see a scenario where 1 year isn’t a problem.
Chinese models are open, available to distill, and they also publish papers about their research. Being one year behind is a skill issue.
I think the most of the money would go to purchase hardware, But I hope they can start making adequate compensation for AI engineers and researchers to move forward.
My feeling is that its a difference in how funding works in different places. The USA will go all in with the populations pensions on a gamble, the Chinese subsidize. This way of operating is typical for the EU.
> Being one year behind is a skill issue
E.g. if you are an AI researcher in Europe you can just go to USA and make generational wealth. This is not a criticism of the EU model, but rather insane American capex effectively monopolizing.
> I think the most of the money would go to purchase hardware,
> > I think the most of the money would go to purchase hardware,
> So it's not a skill issue?
I meant by offering sovereign cloud/inference, not for training. But even if it was for training, Chinese labs have limited supply of GPUs, look what they've done. So it is a skill issue.
Also to clarify, I didn't mean European engineers' skills, I meant "you get what you pay for" as a company, that's why I hope they start offering better compensation to retain talent.
Handing in a photo-copy of a photo-copy of a photo-copy of a photo-copy of a photo-copy of a photo-copy of someone else's homework might work once or twice, but it is no way to run a (non grift) business.
Its not a grift if it works. If a Chinese corp copies at 90% quality at 50% price within a year it means your American business model was never a good one for competing in international markets. The real grift arguably is pretending that it was.
there was "H" at paris at the time, they raised 200M or so, but never got so much visibility. don't know at which stage they are now or if they accomplished something
Wayland should expired, Wayland is architectural failure plagued with bad design decisions, xlibre is gaining momentum and it is not broken piece of a pile
Lol, we don't talk about xlibre here. Announcments and reviews are banned.
Waiting for dang to argue the 7th time that is not somehow his personal decision to ban xlibre that he's set up a system where a couple of nazis with automated scripts can hide an article which requires his manual attention to put (which he doesn't do)
reply