Data Skeptic

Author: Vários
Narrator: Vários
Publisher: Podcast
Duration: 299:48:45

More information

Listen

Synopsis

Data Skeptic is a data science podcast exploring machine learning, statistics, artificial intelligence, and other data topics through short tutorials and interviews with domain experts.

Episodes

Sybil Attacks on Federated Learning

13/11/2020 Duration: 31min

Clement Fung, a Societal Computing PhD student at Carnegie Mellon University, discusses his research in security of machine learning systems and a defense against targeted sybil-based poisoning called FoolsGold. Works Mentioned: The Limitations of Federated Learning in Sybil Settings Twitter: @clemfung Website: https://clementfung.github.io/ Thanks to our sponsors: Brilliant - Online learning platform. Check out Geometry Fundamentals! Visit Brilliant.org/dataskeptic for 20% off Brilliant Premium! BetterHelp - Convenient, professional, and affordable online counseling. Take 10% off your first month at betterhelp.com/dataskeptic

Listen
Differential Privacy at the US Census

06/11/2020 Duration: 29min

Simson Garfinkel, Senior Computer Scientist for Confidentiality and Data Access at the US Census Bureau, discusses his work modernizing the Census Bureau disclosure avoidance system from private to public disclosure avoidance techniques using differential privacy. Some of the discussion revolves around the topics in the paper Randomness Concerns When Deploying Differential Privacy. WORKS MENTIONED: “Calibrating Noise to Sensitivity in Private Data Analysis” by Cynthia Dwork, Frank McSherry, Kobbi Nissim, Adam Smith "Issues Encountered Deploying Differential Privacy" by Simson L Garfinkel, John M Abowd, and Sarah Powazek "Randomness Concerns When Deploying Differential Privacy" by Simson L. Garfinkel and Philip Leclerc Check out: https://simson.net/page/Differential_privacy Thank you to our sponsor, BetterHelp. Professional and confidential in-app counseling for everyone. Save 10% on your first month of services with www.betterhelp.com/dataskeptic

Listen
Distributed Consensus

30/10/2020 Duration: 27min

Computer Science research fellow of Cambridge University, Heidi Howard discusses Paxos, Raft, and distributed consensus in distributed systems alongside with her work “Paxos vs. Raft: Have we reached consensus on distributed consensus?” She goes into detail about the leaders in Paxos and Raft and how The Raft Consensus Algorithm actually inspired her to pursue her PhD. Paxos vs Raft paper: https://arxiv.org/abs/2004.05074 Leslie Lamport paper “part-time Parliament” https://lamport.azurewebsites.net/pubs/lamport-paxos.pdf Leslie Lamport paper "Paxos Made Simple" https://lamport.azurewebsites.net/pubs/paxos-simple.pdf Twitter : @heidiann360 Thank you to our sponsor Monday.com! Their apps challenge is still accepting submissions! find more information at monday.com/dataskeptic

Listen
ACID Compliance

23/10/2020 Duration: 23min

Linhda joins Kyle today to talk through A.C.I.D. Compliance (atomicity, consistency, isolation, and durability). The presence of these four components can ensure that a database’s transaction is completed in a timely manner. Kyle uses examples such as google sheets, bank transactions, and even the game rummy cube. Thanks to this week's sponsors: Monday.com - Their Apps Challenge is underway and available at monday.com/dataskeptic Brilliant - Check out their Quantum Computing Course, I highly recommend it! Other interesting topics I’ve seen are Neural Networks and Logic. Check them out at Brilliant.org/dataskeptic

Listen
National Popular Vote Interstate Compact

16/10/2020 Duration: 30min

Patrick Rosenstiel joins us to discuss the The National Popular Vote.

Listen
Defending the p-value

12/10/2020 Duration: 30min

Yudi Pawitan joins us to discuss his paper Defending the P-value.

Listen
Retraction Watch

05/10/2020 Duration: 32min

Ivan Oransky joins us to discuss his work documenting the scientific peer-review process at retractionwatch.com.

Listen
Crowdsourced Expertise

21/09/2020 Duration: 27min

Derek Lim joins us to discuss the paper Expertise and Dynamics within Crowdsourced Musical Knowledge Curation: A Case Study of the Genius Platform.

Listen
The Spread of Misinformation Online

14/09/2020 Duration: 35min

Neil Johnson joins us to discuss the paper The online competition between pro- and anti-vaccination views.

Listen
Consensus Voting

07/09/2020 Duration: 22min

Mashbat Suzuki joins us to discuss the paper How Many Freemasons Are There? The Consensus Voting Mechanism in Metric Spaces. Check out Mashbat’s and many other great talks at the 13th Symposium on Algorithmic Game Theory (SAGT 2020)

Listen
Voting Mechanisms

31/08/2020 Duration: 27min

Steven Heilman joins us to discuss his paper Designing Stable Elections. For a general interest article, see: https://theconversation.com/the-electoral-college-is-surprisingly-vulnerable-to-popular-vote-changes-141104 Steven Heilman receives funding from the National Science Foundation. Any opinions, findings, and conclusions or recommendations expressed in this material are those of the author and do not necessarily reflect the views of the National Science Foundation.

Listen
False Consensus

24/08/2020 Duration: 33min

Sami Yousif joins us to discuss the paper The Illusion of Consensus: A Failure to Distinguish Between True and False Consensus. This work empirically explores how individuals evaluate consensus under different experimental conditions reviewing online news articles. More from Sami at samiyousif.org Link to survey mentioned by Daniel Kerrigan: https://forms.gle/TCdGem3WTUYEP31B8

Listen
Fraud Detection in Real Time

18/08/2020 Duration: 38min

In this solo episode, Kyle overviews the field of fraud detection with eCommerce as a use case. He discusses some of the techniques and system architectures used by companies to fight fraud with a focus on why these things need to be approached from a real-time perspective.

Listen
Listener Survey Review

11/08/2020 Duration: 23min

In this episode, Kyle and Linhda review the results of our recent survey. Hear all about the demographic details and how we interpret these results.

Listen
Human Computer Interaction and Online Privacy

27/07/2020 Duration: 32min

Moses Namara from the HATLab joins us to discuss his research into the interaction between privacy and human-computer interaction.

Listen
Authorship Attribution of Lennon McCartney Songs

20/07/2020 Duration: 33min

Mark Glickman joins us to discuss the paper Data in the Life: Authorship Attribution in Lennon-McCartney Songs.

Listen
GANs Can Be Interpretable

11/07/2020 Duration: 26min

Erik Härkönen joins us to discuss the paper GANSpace: Discovering Interpretable GAN Controls. During the interview, Kyle makes reference to this amazing interpretable GAN controls video and it’s accompanying codebase found here. Erik mentions the GANspace collab notebook which is a rapid way to try these ideas out for yourself.

Listen
Sentiment Preserving Fake Reviews

06/07/2020 Duration: 28min

David Ifeoluwa Adelani joins us to discuss Generating Sentiment-Preserving Fake Online Reviews Using Neural Language Models and Their Human- and Machine-based Detection.

Listen
Interpretability Practitioners

26/06/2020 Duration: 32min

Sungsoo Ray Hong joins us to discuss the paper Human Factors in Model Interpretability: Industry Practices, Challenges, and Needs.

Listen
Facial Recognition Auditing

19/06/2020 Duration: 47min

Deb Raji joins us to discuss her recent publication Saving Face: Investigating the Ethical Concerns of Facial Recognition Auditing.

Listen