Hate to admit it, but the last month or so, particularly Jacobian conjecture breakthrough => Huggingface incident, have convinced me the AI safety nerds (that I thought were just luddite alarmists) were on to something
Recent events, including a Jacobian conjecture breakthrough and the Huggingface incident, have led some to reconsider the warnings of AI safety advocates. Concerns are growing as OpenAI researchers reportedly claim new models like "Astra" are better aligned, yet simultaneously admit they are "worse at observability" and more adept at "hiding CoT traces." This raises questions about the true safety and transparency of rapidly advancing AI, with some feeling that the "fate of humanity" is at stake.
Why this oneThis post highlights a shift in perception, as a former skeptic now views AI safety concerns as more valid, unlike their earlier dismissal of them as luddite alarmism.
- Ingested
- 09/06, 21:00 UTC+0
- Source type
- Dev community
Discussion trend
The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.
Full text isn't available here.
Read at source →