Wealth could be invested (impacts those borrowing), used on consumables (impacts businesses), or hidden under the mattress or equivalent (reduces the supply of money so acts deflationary).
I would argue that those 3 are the main downstream effects of having wealth -- and they all impact people at large enough values.
This likely means those consuming the outputs of Mechanical Turk don't have a good way to measure the value (aka quality) of the outputs.
If they did - then they shouldn't care whether it's a human or a LLM. And if it's a LLM - then the cost will roughly correlate to the MIN(cost of the LLM, cost of a human) to do the task.
I think the "state of the art" of measuring the quality of outputs was to send the same task to multiple "agents" and only accept answers if over a certain amount agree. With some human review and reputation scoring sprinkled on top. It was a while since I was in this field though
This approach does work when there's a clear answer but what about tasks where the correct answer is multi-modal? Incentivizing agreement works only for tasks where there's clear answers.
The problem is bigger. Outside of coding, there is no real way to reinforce a model with pass/fail cycles until it stops hallucinating. This is why customer service uses will always have a problem. This compounds as you chain agents together.
It's like the speed of light - to get to that point, you need exponentially more energy, and you will never ever get there.
> then they shouldn't care whether it's a human or a LLM.
I imagine that the whole point of posting a task to Mechanical Turk nowadays is that you want it to be completed by humans. Either because you are after the small discrepancy between AI and human performance, or because humans are the object of your investigation.
> If you're a software engineer who cares about health, and have been sitting on the sidelines till now, I think the next few years are a really interesting time to make a contribution.
do you have thoughts on what that looks like? what does the hiring landscape look like?
There is so much to be done and it is not going to be solved by any one thing. That in mind, something like an independent third party patient advocates / advisors would be great. It is hard to be an informed patient these days, Navigating insurance and paperwork is tough and getting second opinions is sometimes had.
Fascinating. I feel like smart person + data (context such as health records, labs, insurance docs, etc.) + LLM -> some helpful tips for navigating the process. However, evaluating the tips on whether they are helpful or not would be hard for a smart person without the expertise.
Or put another way - how does a smart lay person evaluate the independent advocates? Do incentives align?
Right now lots of good health companies hiring, including my own (Empirical). https://www.workatastartup.com/ lets you filter YC companies to see health startups, by stage and location (or remote).
Bayesian statistics to the rescue! You hear about the things that (can) go wrong a lot more, than the babies that are born without an issue. Basically, the opposite of survivorship bias? Anywho, millions of babies are born each day. Only 3-4% are born with a major birth defect or severe health issue (in the US) [0]. Now, a pregnant woman will undergo a routine screening, such as non-invasive prenatal test (NIPT) or a detailed anatomy ultrasound. With a positive test result, the probability of the baby being born healthy goes up to 99%+. Even with a negative test, the probability of the baby being healthy only drops to 80-50%, and then further testing will dial those probabilities in even more.
All this to say, all will be good – don't let the rare anecdotes get to you.
I used to use the docker + homebridge route but it became tedious to maintain.
Instead, I connected it via the Google Home integration (requires an Insights plan) and then use my existing Starling Home hub to access it via HomeKit. This seems to be more reliable and less work than before.
As an alum I've observed both:
* kids who partied hard first semester since grades didn't matter only to struggle second semester
* kids who benefited from this transition because they didn't go to Hotchkiss
I ended up in both buckets and I have mixed feelings about this since it helped and hurt me.
In the long run - it really didn't matter.