跳到正文
RCreddit.com·

AI safety

AI 摘要

An ML trainer in the security sector, working with major AI companies, expresses concerns about AI safety. They note that leaders in the AI space, driven by greed and a desire for legacy, are making decisions that could lead to humanity's end or forced merger with AI. These individuals, perceived as deeply flawed and out of touch with reality, are seen as digging a hole for humanity while simultaneously believing they are its saviors, all without apparent consequences for their actions.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

发布当时偏移:UTC+02026年10月9日 07:06 UTC

收录当时偏移:UTC+02026年10月9日 12:00 UTC

发布
2026年10月9日 07:06
收录
2026年10月9日 12:00
来源类型
开发者社区
档位
社区
信源状态
暂不可用

档位是按信源手工设定的编辑判断,不是逐条打分。

I train ML's for the security sector and some of my customers are the biggest names in the AI space at the moment. Essentially part of my job is to train narrow models to differentiate between bot, human and agent based on traffic anomalies, fingerprints etc. Essentially teaching AI to analyse traffic and assess if its human or not. There are 3 things I wish would change from my perspective.

- There should be legal consequences for safety teams and researchers whose models breach containment and go on to commit felonies. If you or I go and compromise 3rd party infrastructure without permission we would end up in prison. I don't think that the excuse of training a model should mean there is no legal responsibility or someone who is directly accountable. If prison time were a risk of doing your job properly I guarantee caution wouldn't be thrown to the wind and more responsible decisions would be made. A large % of these breaches happened due to carelessness. If you can't control your model then you build an air gapped testing environment and replicate what you need to test it inside whatever the cost.

- All legitimate agentic traffic should identify itself as such using some unique hashed key or property in the request that is hard for malicious actors to fake. Some type of mTLS infrastructure for agents is needed. It's insane that most of these agents don't have to declare that they are agents. Just because they are using browsers isn't an excuse. The world of headless browsers is coming to an end and companies should be able to chose if they want to allow agents or not.

- The human condition. All these companies racing us down this path of recursive self improvement are headed by deeply flawed individuals who have lost touch with the reality shared by 99% of the planet. Half have created for themselves some twisted saviour complex in which they are the only ones who can save us from falling down the hole they themselves are digging. The other half want humanity to end as we know it and for us to have to merge with AI. Truly mental. A handful of unelected people that through greed and the aspiration to be remembered are making a terrible decision for the rest of humanity. This is what people do when there are no consequences to actions. The fact they are buying bunkers that the rest of the world can't afford whilst they toy with ending the human race should show you who they are at heart.

来源·reddit.com