Hacker Newsnew | past | comments | ask | show | jobs | submit | comboy's commentslogin

They are not cracking EUV, they are finding ways that actually scale. Chips can be much cheaper than they are currently. I mean yes, obviously we cannot be confident ntil it happens, but they are already making 5nm with DUV.

When China leads chip making the rest of the world is just vasal states.


Going to higher densities with DUV needs high multi patterning which is very inefficient and ends up costing more than if they could do EUV with fewer patterns.

It's well known tech that is used globally, just not to that degree as yield and throughput goes way down


That's a sharp observation and you're hitting on something most people never even realize.

Whoa, I'm exactly the opposite.

Almost as if evaluating models is mostly astrology as this point.

Well, if you have a clearly defined task it's easy. When using them for my pipeline of writing explanations for Chinese words I have clear ranking, for example - opus 5.5 clearly better than opus 5 at writing and knowing details, annoying nit picker when it comes to finding errors (high accuracy, low usefulness) all in repeatable numbers on different datasets. The problem is that these models are most useful when you are facing a new task that you haven't encountered before. And yup then it's astrology.

It seems to me that often experts from some field will think less of other experts, basically because they have built a different understanding framework. So they both may be equally competent but perceive the other as less competent, and that is just based on the material, excluding some ego stuff.

what character prediction rates are you getting on some unseen datasets?

The held-out scores reported in the Readme IS the unseen dataset.

These are not successful prediction rate per char though.

I got you, will add later to the repository.

Thanks, it seems like a nice way to compare effectiveness of different non-typical methods which are not yet capable of some more ambitious benchmarks.

There's a console and you can track your command! Very cool. Not sure if 3D is needed but very cool nonetheless, It might be that some things are missing (I don't know redis internals though) like some encoding step when getting data back to client, managing client connections and their state (multi, watch) etc.

I agree 3d is not really needed here, it good as an educational thing though.

Awesome dataset. Things are changing. In Chinese-learning community I found that surprising amount of people are creating their own tools without any intent to publish them, it just became easier to create your own thing that to dig out the good stuff from tons of tools already available.


Alignment is a myth. Safety of whom? Humanity couldn't agree on common set of values for thousands of years and we're not gonna suddenly do that in the next ten.


Safety of humans!!! Simple things like not getting killed or enslaved. We could start there...


But what if I want certain other humans to get killed?


Then we should still prioritize the safety of humans


What if I want to smoke cigarettes? Or sell tobacco I grew artisinally to enthusiast tobacco smokers?


Which ones?


Which ones? Because many humans kill other humans rationalizing it by safety of other humans.

I mean I know it seems simple, let's just be excellent to each other. Christianity got pretty far on a decent basic set of values. But it's never simple[1]

1. All the history books


Surely all the AI companies working with the US Department of War shows this is nonsense though? Even if they have accepted Anthropic’s red line of no autonomous lethal weapons, which seems to be the strictest anyone tried to impose, that’s still leaving tonnes of room where they intend AI to help target and kill humans.


But Thiel wants people enslaved and Musk wants then killed. Altman wants them "obsolete" which means desolation.

AfD wants people dead. Right wing men wants women without rights and docile. I could go on ...


The atomic bombings of Japan killed hundreds of thousands of people but most likely "saved" millions.

What should the AI do when asked if it should nuke a country?


Run and present the numbers, then defer.

This is not rocket science.


That wasn't the question.


Yes, it was:

> What should the AI do when asked if it should nuke a country?


Have you tried turning it on and off? More seriously try to remove agent memory or make it refactor it or go through the docs and look for inconsistencies.

But I still don't know how people work with codex when it keeps resetting the context so often, I mean it's surprisingly good at making notes to itself and can follow a long task, but if there are multiple instructions it sometimes forgets some of them.


I just did that and turned on max, let's see if that thing is more of a force!


I'm curious how it went.


26% weekly limit left, 21+ hours strong and still going. I can see the progress though, but I am not confident it will finish before weekly limit is done for (ChatGPT Pro 20x subscription). Since I don't really use LLMs for anything else, except pestering gemini for stupid questions instead of straight up googling it, that's ok. I just wish it would either finish or let me hand it over to it once again whence limit resets.


Perhaps it's better to ponder whether or not it contains information useful to you / is a pleasure to read?


Maybe I’m a Luddite but I find AI prose completely tiresome to read, and its presence in an article forces me to be skeptical about whatever the author wrote. The latter was obviously a problem pre-AI but these tools simply make it much easier to tell stories, narratives, etc. that aren’t yours.


AI-prose may be vetted, but there is no way for a reader to know that for certain as much AI-prose is not.

As such, you cannot be sure if the writing contains correct and accurate information at all, even with modern models.


> you cannot be sure if the writing contains correct and accurate information at all

how is that related to LLMs?


Due to their past, LLMs need to prove themselves as correct and accurate. Bloggers have already gone through this gauntlet and have proven to be more likely to be correct and accurate.

Neither are a guarantee.


Yeah, and there is no pleasure reading this. It's way too long and verbose for what it's trying to say. I'd much rather read the prompt, in bullet points or whatever format they were provided before inflating a few example pictures and notes into a short novel


How is this criticism not just ad hominem slop? You could have said "It's way too long and verbose for what it's trying to say". The writer is orthogonal.


The idea that the writer is orthogonal is WILD


Because you think you're always going to be able to tell the difference between AI writing and human writing? That's WILD!


The heavy usage of LLM writing as is the case here is a strong signal it doesn't and isn't.

I won't waste my time on a person who didn't bother to express themselves.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: