Text classification using... gzip??
Jess Peck on the dangers of AI optimism
Zoom has updated its ToS to allow training AI on user content without the ability to opt-out
Ines Montani and Matthew Honnibal on supervised learning and the ramifications of poor data
Tim Berners-Lee shared the first proposal for World Wide Web on this day in 1989
Quanta's 2022 in Review
Maybe there are some good use cases for ChatGPT after all...
Silly is a Python library for producing silly test data
How to find bad labels in text classification problems using Jupyter and Prodigy
What's the least viewed Wikipedia article? Colin Morris tried to find out.
A MiniDisc Player can now support full data transfer thanks to this hack
The pros and cons of language models searching the Web for its data
Beating bias in AI datasets is more than just data diversity—it's a balance of data and training
data / ethics / MIT (Massachusetts Institute of Technology) / neural networks / neuroscience / research