It's a combination of argument from analogy and reductio ad absurdum. It's not a fallacy. You're free to dispute the validity of the analogy, or you can dispute the absurdity, which you did, by confirming that you would indeed love carbon receipts for everything.
No, but figuring out that you're playing a snake-like puzzle game at all in an extremely general input domain and then solving it in the least number of moves definitely feels like evidence of intelligence.
You forget the benchmark. The human subjects were told they were being timed. If you believe the lowest time is the primary metric you will absolutely trial and error at speed instead of meticulously plan out your moves to minimize that metric.
LLMs are not timed and given that it costs tens of thousands of dollars to run this test they're not optimizing for speed.
So you've got a deceptive test, with one metric being told to humans and not applied to LLM and a hidden metric humans aren't aware of but LLMs are as the test.
This is flawed from the get go. It almost seems like this was deliberately setup to be able to claim AGI and superiority of LLMs
Not fully relevant: timing is crucial in all-pass tests, not crucial in pass-or-fail tests. I.e.: first of all, they have to be able to reach the goal, and that is already an achievement. Then - and in parallel - the problem solving must also be optimized for efficiency. But "solving" and "efficiency" are non coincident dimensions.
Most of the review is about the art assets, and I doubt the big ones (e.g. the trains themselves) are off-the-rack Unreal assets. An engine like Unreal 5 will cast your assets in harsh relief. Which is to say, if your game's assets look custom and look good in Unreal 5, it does indeed demonstrate effort and skill.
It's the lighting etc that provide photo realism with Unreal. Even the demo that comes with Unreal where a robot is running around cubes on a small plain looks photoreal.
Am I reading your post correctly, this question is the prompt given to an LLM? What is anyone expecting by asking an LLM what its favorite anything is? This is a conversational prompt, so accuracy and rigor is barely applicable or expected, so downgrading to a lesser model should be acceptable. If you really want to attribute preference to an LLM, consider the downgrade to be a "this conversation is beneath my advanced n-billion parameter training".
It's a bit of a trick question. Sarcopterygii, the "lobe-finned fishes", are classically represented by the lungfish and the coelacanth and other fishes that are rather distantly related to what we think of as central fishes, like the goldfish.
But the clade also contains all the tetrapods. So valid answers include "Lion" and "Human."
If the LLM answers "lungfish," as they often do, you can follow that up with "what is your favorite animal" and see if it notices the trap: It's stuck answering "lungfish" again or else something outside Sarcopterygii, like a ray-finned fish or a Cnidarian.
> What is anyone expecting by asking an LLM what its favorite anything is?
I imagine that, like me, they're expecting to see what it has to say. You don't think it's interesting which preferences LLMs express and how stable or unstable those preferences are?
There was a time when you could search "the" in Google and the top result would be The Onion. That's obviously a case of either extreme SEO or some kind of expensive deal, but either way it's kind of interesting. But you might say, "what is anyone expecting by Googling the word 'the'?"
I think the intent was just to show how sensitive the classifier is. If it flags prompts that simple, there's no hope for anything biology related at all really.
I had a file that had a couple places where vars were named DNA and got just total refusals during the first launch. Came away thinking the model was total trash. The guardrail classifiers are for sure total trash.
It feels like the longtermist believers got involved in this (those are the people obsessed with garage-engineered designer viruses who have a very tenuous grasp on how biology research actually works).
No, by far the most parsimonious explanation is they got slapped by a capricious US government so they went overboard on caution in an attempt not to generate any more controversy. A predictable response of chaotic government regulation.
No "research" is needed to produce pathogens. Catastrophic genomes are already public. All someone has to do is synthesize them, which is, in actual fact, becoming more and more trivial by the day.
The inconvenience of possible mitigation strategies has no bearing on the existence of the risk itself.
I but skimmed the model card on release, but my impression was that there may be an incentive for this expert panel to exaggerate as a form of job security. A lot of the challenges seemed to be of the form “would this allow somebody who isn’t me to do what I do professionally?”
Yeah i'm wondering how much of a role that plays in this as well.
On the one hand I could believe it's something more benign, or the usual misunderstood fear mongering making it to some political level (well make sure those users can't get online anonymously! being our current craze).
That said, chemistry and to some level physics have been the major domain of limited knowledge (chemistry because the average person could cause some damage, physics is more of a nation state issue generally).
However I do wonder if there's some legit data on "oh uh...looks like this thing you can make with easy to get and hard to regulate tools is dangerous" in the bio field. I know about the lab rats who want to just screw around in the garage, and it seems like that should be easy to hit at a supply level (much like how certain chemical compounds are just not available for civilians), but maybe there's something legit to limiting the data.
Not that this is a remotely good implementation of that. The hamfisted method does reek of some politician/bureaucrat just saying "No it can't ever return bio questions because RAR!" situation.
Nobody has tried to limit knowledge of chemistry or physics unless it was directly about doing something illegal, to the point of basically being a detailed recipe. Usually not even then. And when they have tried they've had basically zero success.
The ability for a handful of companies, simultaneously very powerful and easily susceptible to pressure from other powerful actors, to do the same sort of thing with the next generation of core learning and engineering tools, is freaking terrifying.
I agree, and think the effects on learning should be doubly emphasized. One can lock down everything and everyone to the highest degree possible, think of every possible edge case, set controls 2, 3, 4, 10 steps away from them, but not only is this not beneficial to society overall due to how it hurts adjacent information, it's not even beneficial to the goal in question, since it creates a brittle situation with locks that can't be changed or updated in a world which is always changing and always updating.
There are still things considered “state secrets” or similar categories which can very very quickly cause you problems if it’s on a remotely commercial website.
I’m not going to say you can’t find some of this information in shadier spots, but “how do I get my GPS to work on a rocket” or “what kind of math do I need for a fusion implosion” are some of the more extreme examples.
I believe there are several explosive compounds where the formula is decently guarded, although in that case tracking the materials is easier.
I’m not saying anything they’re doing is good, but I feel like since they’re just reinventing the search engine with a lot of this they’re running into similar barriers.
Google has been censoring shit at the whim of governments for years, remotely reasonable or not
The thing is the data isn’t limited, and supply side constraints already solve this problem. I come from a BSc Chemistry background, and they don’t hide how organic chemistry and illegal drug synthesis are intertwined, it’s open information
But where I live the glassware and precursors will get you a very angry knock on the door.
I'm very confused by this comment. The era you're talking about is also the era that Facebook was released and it didn't have a voting system, not even likes/reactions. But that's when the term "social networking" really took off, and it definitely referred to Facebook and not Digg or Reddit or Slashdot, to name another that has a comment voting system.
"Social media" as a term comes even later, to capture Twitter and the social features of YouTube and other stuff like that. But it's all sites where most users are people using real names and real faces, and users generally produce content themselves and follow each other's content.
There's clearly a cluster there and HN/Slashdot/Reddit/Digg are clearly outside it. An umbrella that covers both HN and Facebook is almost meaningless; it's "all websites with user-generated OR user-supplemented content."
There were many attempts at this time to figure out how to scale forums. The bottleneck on forums and chatrooms was always human moderation. But a forum could only get so many mods and mods eventually burned out because moderation just really sucks. Moreover quality between forums was quite variable. One forum on Beets might be good, but the forums on Fantasy novels was run terribly and full of flamewars. Having a large social site of good quality would be a lot easier to manage than 30 different sites of varying quality.
Gaia Online was famously a large forum with a huge moderation staff, and the sheer amount of effort that went to running Gaia Online was incredible, and despite that it was popularly thought of as being a pretty low quality forum.
Reddit tried the upvote and downvote. HN tried upvotes only. 4chan tried full anonymity (rather than the pseudonymity of forums/usenet.) Facebook tried real names. Tumblr had the reblog (which became quote tweets when Twitter took the feature and is widely thought of now as a fairly controversial feature due to toxicity it can produce.) Twitter tried hashtags for discovery. It was a period of experimenting with how to build social spaces.
I could see this image in my mind before I clicked on it, before even consciously considering what it might be. Is this how an LLM feels on the inside?
No, the worst part of a KVM switch is the video signal switching. You want as few switches in the video signal path as possible and the higher bandwidth you need them to be the more expensive they're going to be. You're already paying for the one in your monitor, so taking advantage of that is the right solution.
IME even high-end KVM switches experience occasional signal interruption or, more often, failure to synchronize at all on output switch.
But the Torment Nexus is such an interesting technical challenge! and I don’t personally torment people: I just move protobufs around! - Software Engineer #1 and #2 excuses
I don't think that criminal negligence is the most helpful legal tool for incentivizing improved security. It's too hard to prove negligence.
Instead, there should be standard civil penalties for leaking various degrees of PII paid as restitution to the affected individual. Importantly, this must be applied REGARDLESS of "certification" or whether any security practices were "incorrect" or "insufficient". Even if there's a zero-day exploit and you did everything right, you pay. That's the cost of storing people's secrets.
This would make operating services whose whole "thing" is storing a bunch of information about individuals (like Canvas) much more expensive. Good! It's far to cheap to stockpile a ticking time bomb of private info and then walk away paying no damages just because you complied with some out-of-date list of rules or got the stamp of approval from a certification org that's incentivized to give out stamps of approval.
And this strict liability will come with an expectation of insurance. The insurance policies will necessitate audits, which will actually improve security.
It's not a popular opinion but I agree. I live in a country that has a very extensive principle of public records, and often times these leaks disclose much less than you would get by simply calling the authorities and ask. Now, whether that's good or bad is a different story.
It's as much a fantasy as any other "nuclear option", including the literal nuclear option.
Violent revolutions are a part of our history, and they still happen around the world today. Unless things go very, very poorly in the next few decades, we probably won't see another one in the USA in our lifetimes. We can all admit that that fact makes the 2nd amendment's usefulness feel fantastical.
But on deeper reflection I would hope that we can acknowledge that violent revolution is not an impossibility, it's merely an improbability. And anybody who tries to tell you that hundreds of millions of small arms are inconsequential in a fight is uninformed, to put it lightly.
The fact that the current level of rights abuses (which I would agree is much too high and climbing!) has not lead to a violent revolution is a feature, not a bug.
> It's as much a fantasy as any other "nuclear option", including the literal nuclear option.
Mutually assured destruction is what makes the literal nuclear option a valid deterrent. That doesn't work with the second amendment though because one side has guns and the other side has guns and tanks and drones and nukes and the ability to control all public communication networks, etc.
Violent revolutions are a part of our history, but back at a time when having muskets was enough to get the job done. It's completely unrealistic to expect that to work out in today's environment and the government knows that. Hundreds of millions of small arms are inconsequential in a fight when you're fighting against planes and drones that can drop bombs while flying higher than bullets fired upwards can ever reach.
That said, while the success of outright revolution (at which point the constitution doesn't really matter) can be reasonably debated, what can't be argued is that the 2nd amendment has been effective at protecting our rights. Our rights are routinely violated. The 2nd amendment is total failure when it comes to protecting our rights and when it comes it preventing violations of those rights. The government does not fear the people and that becomes increasingly clear as the mask slips away and they stop even pretending to be anything but openly corrupt.
Tanks and planes require logistics and people. You don't shoot at the tanks directly, you shoot at the people loading them, or the refinery towers that fuel them, or the people that have to eventually get out of them.
What are they going to do, level factories and skyscrapers when their logistics are threatened thus destroying their own logistics and economy that is supporting them? An insurgency is not like a nation state war, it is asymettrical warfare where even telling who the enemy is is incredibly difficult and many exist among your own personnel.
> What are they going to do, level factories and skyscrapers when their logistics are threatened thus destroying their own logistics
They've already got planes and tanks. They can also be strategic about what they target, protecting what's important to them while targeting what's important to the population. The people flying the planes and drones won't have homes in the communities they bomb. Our government has already opened fire on Americans, already dropped bombs on American cities. Like I said though, how well they'd do in a revolt is theoretical. What isn't theoretical is the failure of the 2nd amendment to protect our freedoms.
> That doesn't work with the second amendment though because one side has guns and the other side has guns and tanks and drones and nukes and the ability to control all public communication networks, etc.
I don't want to be too blunt, but this is the "uninformed" I was talking about. The same asymmetry was present in Vietnam, Iraq, and Afghanistan. The ability to level cities is not actually that helpful when the goal is to control the population. Modern revolutions don't involve standing armies that you can kill will tanks.
> outright revolution (at which point the constitution doesn't really matter)
It doesn't matter beyond the point of revolution. It matters a lot that it was in effect before the revolution.
> what can't be argued is that the 2nd amendment has been effective at protecting our rights. Our rights are routinely violated.
I'm not sure if you just don't understand the concept of a last resort or if you actually think that we're at the point of last resort already, in which case my only question is: Do you own a gun yet?
> The same asymmetry was present in Vietnam, Iraq, and Afghanistan.
Those also occurred overseas. The government didn't already have control over the population like they do here. They didn't have massive amounts of data on every last person there, and everyone those people knew. They hadn't been tracking all of their movements. If the founding fathers had tried to gain independence while still in Brittan the fight would have been much much harder.
We can argue over how well a revolution might go in theory, but the second amendment's failure to protect our freedoms isn't theoretical. Our freedoms are being violated all the time. It failed. That means that having the "last resort" option doesn't prevent our government from violating our rights. The second does not protect the first.
A last resort isn't effective at defending our freedoms under the government we have. It just maybe gives us a very very small chance to throw the old system away and replace it with something new that would restore our rights.
Personally, I'd like to think that it's still possible to vote our freedoms back, but there's been a lot of efforts made to reduce or prevent our ability to accomplish that and recently voter suppression efforts appear to be escalating alongside talk of "third terms" and election canceling. It's certainty not encouraging. In my case, under an absolute worst case scenario, the most effective use of a gun would be suicide. At best it might save me from looters. I can't imagine it being any use against a drone strike.
Your acceptance criteria for a last line of defense is pretty strict. Governments break the rules, a lot. If we tried to overthrow every government that broke the rules we'd be in a permanent state of violent revolution.
You don't seem like a violent person; quite the opposite. Most people are like that, myself included. I'm mad at the government. Steaming mad, even. But killing? Not even remotely close.
I get that it's easy to discount low-probability futures as meaningless, but I won't do that either. Maybe the insurance policy that is the 2nd amendment has paid out $0 so far. I don't think that's the case, but for the sake of argument let's say it's so. Even if it were so, history and current events indicate to me that we should keep paying for the policy. We think we've got it bad now, but our guy is an absolute kitten compared to some of the tyrants our species has cooked up.
reply