Thousands of Internal AI Training Datasets, Tools Exposed to Anyone on the Internet
Full article177 words · extracted from 404media.co · click to collapse
Thousands of machine learning tools, including some belonging to large tech companies, are exposed to the open internet, letting anyone interact with them and potentially exposing sensitive data, according to a security researcher who shared their findings with 404 Media.
The news shows that even while companies and researchers barrel ahead with artificial intelligence research, securing those tools sometimes boils down to the same old account security and authentication best practices that apply to other types of accounts.
“In addition to the ML models themselves, the exposed data can include training datasets, hyperparameters, and sometimes even raw data used to build models,” Charan Akiri, the security researcher and a lead security engineer at Reddit, said in a write-up of his research.
This post is for paid members only
Become a paid member for unlimited ad-free access to articles, bonus podcast content, and more.
Sign up for free access to this post
Free members get access to posts like this one along with an email round-up of our week's stories.
Already have an account? Sign in
Text extracted automatically; images, tables and formatting may be missing. Original: https://www.404media.co/thousands-of-internal-ai-training-datasets-tools-exposed-to-anyone-on-the-internet/