muckrAIkers – Podcast

Afleveringen

DeepSeek Minisode
10 feb· muckrAIkers
DeepSeek R1 has taken the world by storm, causing a stock market crash and prompting further calls for export controls within the US. Since this story is still very much in development, with follow-up investigations and calls for governance being released almost daily, we thought it best to hold of for a little while longer to be able to tell the whole story. Nonetheless, it's a big story, so we provide a brief overview of all that's out there so far.
(00:00) - Recording date(00:04) - Intro(00:37) - DeepSeek drop and reactions(04:27) - Export controls(08:05) - Skepticism and uncertainty(14:12) - Outro

LinksDeepSeek websiteDeepSeek paperReuters article - What is DeepSeek and why is it disrupting the AI sector?
Fallout coverage
The Verge article - OpenAI has evidence that its models helped train China’s DeepSeekThe Signal article - Nvidia loses nearly $600 billion in DeepSeek crashCNN article - US lawmakers want to ban DeepSeek from government devicesFortune article - Meta is reportedly scrambling ‘war rooms’ of engineers to figure out how DeepSeek’s AI is beating everyone else at a fraction of the priceDario Amodei's blogpost - On DeepSeek and Export ControlsSemiAnalysis article - DeepSeek DebatesArs Technica article - Microsoft now hosts AI model accused of copying OpenAI dataWiz Blogpost - Wiz Research Uncovers Exposed DeepSeek Database Leaking Sensitive Information, Including Chat History
Investigations into "reasoning"
Blogpost - There May Not be Aha Moment in R1-Zero-like Training — A Pilot StudyPreprint - s1: Simple test-time scalingPreprint - LIMO: Less is More for ReasoningBlogpost - Reasoning ReflectionsPreprint - Token-Hungry, Yet Precise: DeepSeek R1 Highlights the Need for Multi-Step Reasoning Over Speed in MATH
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
Understanding AI World Models w/ Chris Canal
27 jan· muckrAIkers
Chris Canal, co-founder of EquiStamp, joins muckrAIkers as our first ever podcast guest! In this ~3.5 hour interview, we discuss intelligence vs. competencies, the importance of test-time compute, moving goalposts, the orthogonality thesis, and much more.
A seasoned software developer, Chris started EquiStamp as a way to improve our current understanding of model failure modes and capabilities in late 2023. Now a key contractor for METR, EquiStamp evaluates the next generation of LLMs from frontier model developers like OpenAI and Anthropic.
EquiStamp is hiring, so if you're a software developer interested in a fully remote opportunity with flexible working hours, join the EquiStamp Discord server and message Chris directly; oh, and let him know muckrAIkers sent you!
(00:00) - Recording date(00:05) - Intro(00:29) - Hot off the press(02:17) - Introducing Chris Canal(19:12) - World/risk models(35:21) - Competencies + decision making power(42:09) - Breaking models down(01:05:06) - Timelines, test time compute(01:19:17) - Moving goalposts(01:26:34) - Risk management pre-AGI(01:46:32) - Happy endings(01:55:50) - Causal chains(02:04:49) - Appetite for democracy(02:20:06) - Tech-frame based fallacies(02:39:56) - Bringing back real capitalism(02:45:23) - Orthogonality Thesis (03:04:31) - Why we do this(03:15:36) - Equistamp!

Links
EquiStampChris's TwitterMETR Paper - RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human expertsAll Trades article - Learning from History: Preventing AGI Existential Risks through Policy by Chris CanalBetter Systems article - The Omega Protocol: Another Manhattan Project
Superintelligence & Commentary
Wikipedia article - Superintelligence: Paths, Dangers, Strategies by Nick BostromReflective Altruism article - Against the singularity hypothesis (Part 5: Bostrom on the singularity)Into AI Safety Interview - Scaling Democracy w/ Dr. Igor Krawczuk
Referenced Sources
Book - Man-made Catastrophes and Risk Information Concealment: Case Studies of Major Disasters and Human FallibilityArtificial Intelligence Paper - Reward is EnoughWikipedia article - Capital and Ideology by Thomas PikettyWikipedia article - Pantheon
LeCun on AGI
"Won't Happen" - Time article - Meta’s AI Chief Yann LeCun on AGI, Open-Source, and AI Risk"But if it does, it'll be my research agenda latent state models, which I happen to research" - Meta Platforms Blogpost - I-JEPA: The first AI model based on Yann LeCun’s vision for more human-like AI
Other Sources
Stanford CS Senior Project - Timing Attacks on Prompt Caching in Language Model APIsTechCrunch article - AI researcher François Chollet founds a new AI lab focused on AGIWhite House Fact Sheet - Ensuring U.S. Security and Economic Strength in the Age of Artificial IntelligenceNew York Post article - Bay Area lawyer drops Meta as client over CEO Mark Zuckerberg’s ‘toxic masculinity and Neo-Nazi madness’OpenEdition Academic Review of Thomas PikettyNeural Processing Letters Paper - A Survey of Encoding Techniques for Signal Processing in Spiking Neural NetworksBFI Working Paper - Do Financial Concerns Make Workers Less Productive?No Mercy/No Malice article - How to Survive the Next Four Years by Scott Galloway
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
Zijn er afleveringen die ontbreken?

Klik hier om de feed te vernieuwen.
NeurIPS 2024 Wrapped 🌯
30 dec 2024· muckrAIkers
What happens when you bring over 15,000 machine learning nerds to one city? If your guess didn't include racism, sabotage and scandal, belated epiphanies, a spicy SoLaR panel, and many fantastic research papers, you wouldn't have captured my experience. In this episode we discuss the drama and takeaways from NeurIPS 2024.
Posters available at time of episode preparation can be found on the episode webpage.
EPISODE RECORDED 2024.12.22
(00:00) - Recording date(00:05) - Intro(00:44) - Obligatory mentions(01:54) - SoLaR panel(18:43) - Test of Time(24:17) - And now: science!(28:53) - Downsides of benchmarks(41:39) - Improving the science of ML(53:07) - Performativity(57:33) - NopenAI and Nanthropic(01:09:35) - Fun/interesting papers(01:13:12) - Initial takes on o3(01:18:12) - WorkArena(01:25:00) - Outro

Links
Note: many workshop papers had not yet been published to arXiv as of preparing this episode, the OpenReview submission page is provided in these cases.
NeurIPS statement on inclusivityCTOL Digital Solutions article - NeurIPS 2024 Sparks Controversy: MIT Professor's Remarks Ignite "Racism" Backlash Amid Chinese Researchers’ Triumphs(1/2) NeurIPS Best Paper - Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale PredictionVisual Autoregressive Model report this link now provides a 404 errorDon't worry, here it is on archive.isReuters article - ByteDance seeks $1.1 mln damages from intern in AI breach case, report saysCTOL Digital Solutions article - NeurIPS Award Winner Entangled in ByteDance's AI Sabotage Accusations: The Two Tales of an AI GeniusReddit post on Ilya's talkSoLaR workshop page
Referenced Sources
Harvard Data Science Review article - Data Science at the SingularityPaper - Reward Reports for Reinforcement LearningPaper - It's Not What Machines Can Learn, It's What We Cannot TeachPaper - NeurIPS Reproducibility ProgramPaper - A Metric Learning Reality Check
Improving Datasets, Benchmarks, and Measurements
Tutorial video + slides - Experimental Design and Analysis for AI Researchers (I think you need to have attended NeurIPS to access the recording, but I couldn't find a different version)Paper - BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best PracticesPaper - Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?Paper - A Systematic Review of NeurIPS Dataset Management PracticesPaper - The State of Data Curation at NeurIPS: An Assessment of Dataset Development Practices in the Datasets and Benchmarks TrackPaper - Benchmark Repositories for Better BenchmarkingPaper - Croissant: A Metadata Format for ML-Ready DatasetsPaper - Rethinking the Evaluation of Out-of-Distribution Detection: A Sorites ParadoxPaper - Evaluating Generative AI Systems is a Social Science Measurement ChallengePaper - Report Cards: Qualitative Evaluation of LLMs
Governance Related
Paper - Towards Data Governance of Frontier AI ModelsPaper - Ways Forward for Global AI Benefit SharingPaper - How do we warn downstream model providers of upstream risks?Unified Model Records toolPaper - Policy Dreamer: Diverse Public Policy Creation via Elicitation and Simulation of Human PreferencesPaper - Monitoring Human Dependence on AI Systems with Reliance DrillsPaper - On the Ethical Considerations of Generative AgentsPaper - GPAI Evaluation Standards Taskforce: Towards Effective AI GovernancePaper - Levels of Autonomy: Liability in the age of AI Agents
Certified Bangers + Useful Tools
Paper - Model Collapse Demystified: The Case of RegressionPaper - Preference Learning Algorithms Do Not Learn Preference RankingsLLM Dataset Inference paper + repodattri paper + repoDeTikZify paper + repo
Fun Benchmarks/Datasets
Paloma paper + datasetRedPajama paper + datasetAssemblage webpageWikiDBs webpageWhodunitBench repoApeBench paper + repoWorkArena++ paper
Other Sources
Paper - The Mirage of Artificial Intelligence Terms of Use Restrictions
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
OpenAI's o1 System Card, Literally Migraine Inducing
23 dec 2024· muckrAIkers
The idea of model cards, which was introduced as a measure to increase transparency and understanding of LLMs, has been perverted into the marketing gimmick characterized by OpenAI's o1 system card. To demonstrate the adversarial stance we believe is necessary to draw meaning from these press-releases-in-disguise, we conduct a close read of the system card. Be warned, there's a lot of muck in this one.
Note: All figures/tables discussed in the podcast can be found on the podcast website at https://kairos.fm/muckraikers/e009/
(00:00) - Recorded 2024.12.08(00:54) - Actual intro(03:00) - System cards vs. academic papers(05:36) - Starting off sus(08:28) - o1.continued(12:23) - Rant #1: figure 1(18:27) - A diamond in the rough(19:41) - Hiding copyright violations(21:29) - Rant #2: Jacob on "hallucinations"(25:55) - More ranting and "hallucination" rate comparison(31:54) - Fairness, bias, and bad science comms(35:41) - System, dev, and user prompt jailbreaking(39:28) - Chain-of-thought and Rao-Blackwellization(44:43) - "Red-teaming"(49:00) - Apollo's bit(51:28) - METR's bit(59:51) - Pass@???(01:04:45) - SWE Verified(01:05:44) - Appendix bias metrics(01:10:17) - The muck and the meaning

Linkso1 system cardOpenAI press release collection - 12 Days of OpenAI

Additional o1 Coverage
NIST + AISI [report] - US AISI and UK AISI Joint Pre-Deployment TestApollo Research's paper - Frontier Models are Capable of In-context SchemingVentureBeat article - OpenAI launches full o1 model with image uploads and analysis, debuts ChatGPT ProThe Atlantic article - The GPT Era Is Already Ending

On Data Labelers
60 Minutes article + video - Labelers training AI say they're overworked, underpaid and exploited by big American tech companiesReflections article - The hidden health dangers of data labeling in AI developmentPrivacy International article = Humans in the AI loop: the data labelers behind some of the most powerful LLMs' training datasets

Chain-of-Thought Papers Cited
Paper - Measuring Faithfulness in Chain-of-Thought ReasoningPaper - Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought PromptingPaper - On the Hardness of Faithful Chain-of-Thought Reasoning in Large Language ModelsPaper - Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Other Mentioned/Relevant Sources
Andy Jones blogpost - Rao-BlackwellizationPaper - Training on the Test Task Confounds Evaluation and EmergencePaper - Best-of-N JailbreakingResearch landing page - SWE BenchCode Competition - Konwinski PrizeLakera game = GandalfKate Crawford's Atlas of AIBlueDot Impact's course - Intro to Transformative AI

Unrelated Developments
Cruz's letter to Merrick GarlandAWS News Blog article - Introducing Amazon Nova foundation models: Frontier intelligence and industry leading price performanceBleepingComputer article - Ultralytics AI model hijacked to infect thousands with cryptominerThe Register article - Microsoft teases Copilot Vision, the AI sidekick that judges your tabsFox Business article - OpenAI CEO Sam Altman looking forward to working with Trump admin, says US must build best AI infrastructure
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
How to Safely Handle Your AGI
2 dec 2024· muckrAIkers
While on the campaign trail, Trump made claims about repealing Biden's Executive Order on AI, but what will actually be changed when he gets into office? We take this opportunity to examine policies being discussed or implemented by leading governments around the world.
(00:00) - Intro(00:29) - Hot off the press(02:59) - Repealing the AI executive order?(11:16) - "Manhattan" for AI(24:33) - EU(30:47) - UK(39:27) - Bengio(44:39) - Comparing EU/UK to USA(45:23) - China(51:12) - Taxes(55:29) - The muck

LinksSFChronicle article - US gathers allies to talk AI safety as Trump's vow to undo Biden's AI policy overshadows their workTrump's Executive Order on AI (the AI governance executive order at home)Biden's Executive Order on AICongressional report brief which advises a "Manhattan Project for AI"
Non-USA
CAIRNE resource collection on CERN for AIUK Frontier AI Taskforce report (2023)International interim report (2024)Bengio's paper - AI and Catastrophic RiskDavidad's Safeguarded AI program at ARIAMIT Technology Review article - Four things to know about China’s new AI rules in 2024GovInsider article - Australia’s national policy for ethical use of AI starts to take shapeFuture of Privacy forum article - The African Union’s Continental AI Strategy: Data Protection and Governance Laws Set to Play a Key Role in AI Regulation
Taxes
Macroeconomic Dynamics paper - Automation, Stagnation, and the Implications of a Robot TaxCESifo paper - AI, Automation, and TaxationGavTax article - Taxation of Artificial Intelligence and Automation
Perplexity Pages
CERN for AI pageChina's AI policy pageSingapore's AI policy pageAI policy in Africa, India, Australia page
Other Sources
Artificial Intelligence Made Simple article - NYT's "AI Outperforms Doctors" Story Is WrongIntel report - Reclaim Your Day: The Impact of AI PCs on ProductivityHeise Online article - Users on AI PCs slower, Intel sees problem in unenlightened usersThe Hacker News article - North Korean Hackers Steal $10M with AI-Driven Scams and Malware on LinkedInFuturism article - Character.AI Is Hosting Pedophile Chatbots That Groom Users Who Say They're UnderageVice article - 'AI Jesus' Is Now Taking Confessions at a Church in SwitzerlandPolitico article - Ted Cruz: Congress 'doesn't know what the hell it's doing' with AI regulationUS Senate Committee on Commerce, Science, and Transportation press release - Sen. Cruz Sounds Alarm Over Industry Role in AI Czar Harris’s Censorship Agenda
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
The End of Scaling?
19 nov 2024· muckrAIkers
Multiple news outlets, including The Information, Bloomberg, and Reuters [see sources] are reporting an "end of scaling" for the current AI paradigm. In this episode we look into these articles, as well as a wide variety of economic forecasting, empirical analysis, and technical papers to understand the validity, and impact of these reports. We also use this as an opportunity to contextualize the realized versus promised fruits of "AI".
(00:23) - Hot off the press(01:49) - The end of scaling(10:50) - "Useful tools" and "agentic" "AI"(17:19) - The end of quantization(25:18) - Hedging(29:41) - The end of upwards mobility(33:12) - How to grow an economy(38:14) - Transformative & disruptive tech(49:19) - Finding the meaning(56:14) - Bursting AI bubble and Trump(01:00:58) - The muck
LinksThe Information article - OpenAI Shifts Strategy as Rate of ‘GPT’ AI Improvements SlowsBloomberg [article] - OpenAI, Google and Anthropic Are Struggling to Build More Advanced AIReuters article - OpenAI and others seek new path to smarter AI as current methods hit limitationsPaper on the end of quantization - Scaling Laws for PrecisionTim Dettmers Tweet on "Scaling Laws for Precision"
Empirical Analysis
WU Vienna paper - Unslicing the pie: AI innovation and the labor share in European regionsIMF paper - The Labor Market Impact of Artificial Intelligence: Evidence from US RegionsNBER paper - Automation, Career Values, and Political PreferencesPew Research Center report - Which U.S. Workers Are More Exposed to AI on Their Jobs?
Forecasting
NBER/Acemoglu paper - The Simple Macroeconomics of AINBER/Acemoglu paper - Harms of AIIMF report - Gen-AI: Artificial Intelligence and the Future of WorkSubmission to Open Philanthropy AI Worldviews Contest - Transformative AGI by 2043 is <1% likely
Externalities and the Bursting Bubble
NBER paper - Bubbles, Rational Expectations and Financial MarketsClayton Christensen lecture capture - Clayton Christensen: Disruptive innovationThe New Republic article - The “Godfather of AI” Predicted I Wouldn’t Have a Job. He Was Wrong.Latent Space article - $2 H100s: How the GPU Rental Bubble Burst
On Productization
Palantir press release on introduction of Claude to US security and defenseArs Technica article - Claude AI to process secret government data through new Palantir dealOpenAI press release on partnering with Condé NastCandid Technology article - Shutterstock and Getty partner with OpenAI and BRIAE2BStripe agentsRobopair
Other Sources
CBS News article - Google AI chatbot responds with a threatening message: "Human … Please die."Biometric Update article - Travelers to EU may be subjected to AI lie detectorTechcrunch article - OpenAI’s tumultuous early years revealed in emails from Musk, Altman, and othersRichard Ngo Tweet on leaving OpenAI
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
US National Security Memorandum on AI, Oct 2024
6 nov 2024· muckrAIkers
October 2024 saw a National Security Memorandum and US framework for using AI in national security contexts. We go through the content so you don't have to, pull out the important bits, and summarize our main takeaways.
(00:48) - The memorandum(06:28) - What the press is saying(10:39) - What's in the text(13:48) - Potential harms(17:32) - Miscellaneous notable stuff(31:11) - What's the US governments take on AI?(45:45) - The civil side - comments on reporting(49:31) - The commenters(01:07:33) - Our final hero(01:10:46) - The muck

LinksUnited States National Security Memorandum on AIFact Sheet on the National Security MemorandumFramework to Advance AI Governance and Risk Management in National Security
Related Media
CAIS Newsletter - AI Safety Newsletter #43NIST report - Artificial Intelligence Risk Management Framework: Generative Artificial Intelligence ProfileACLU press release - ACLU Warns that Biden-Harris Administration Rules on AI in National Security Lack Key ProtectionsWikipedia article - Presidential MemorandumReuters article - White House presses gov't AI use with eye on security, guardrailsForbes article - America’s AI Security Strategy Acknowledges There’s No Stopping AIDefenseScoop article - New White House directive prods DOD, intelligence agencies to move faster adopting AI capabilitiesNYTimes article - Biden Administration Outlines Government ‘Guardrails’ for A.I. ToolsForbes article - 5 Things To Know About The New National Security Memorandum On AI – And What ChatGPT ThinksFederal News Network interview - A look inside the latest White House artificial intelligence memoGovtech article - Reactions Mostly Positive to National Security AI MemoThe Information article - Biden Memo Encourages Military Use of AI
Other Sources
Physical Intelligence press release - π0: Our First Generalist PolicyOpenAI press release - Introducing ChatGPT SearchWhoPoo App!!
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
Understanding Claude 3.5 Sonnet (New)
30 okt 2024· muckrAIkers
Frontier developers continue their war on sane versioning schema to bring us Claude 3.5 Sonnet (New), along with "computer use" capabilities. We discuss not only the new model, but also why Anthropic may have released this model and tool combination now.
(00:00) - Intro(00:22) - Hot off the press(05:03) - Claude 3.5 Sonnet (New) Two 'o' 3000(09:23) - Breaking down "computer use"(13:16) - Our understanding(16:03) - Diverging business models(32:07) - Why has Anthropic chosen this strategy?(43:14) - Changing the frame(48:00) - Polishing the lily
Links
Anthropic press release - Introducing Claude 3.5 Sonnet (New)Model Card Addendum
Other Anthropic Relevant Media
Paper - Sabotage Evaluations for Frontier ModelsAnthropic press release - Anthropic's Updated RSPAlignment Forum blogpost - Anthropic's Updated RSPTweet - Response to scare regarding Anthropic training on user dataAnthropic press release - Developing a computer use modelSimon Willison article - Initial explorations of Anthropic’s new Computer Use capabilityTweet - ARC Prize performanceThe Information article - Anthropic Has Floated $40 Billion Valuation in Funding Talks
Other Sources
LWN.net article - OSI readies controversial Open AI definitionNational Security MemorandumFramework to Advance AI Governance and Risk Management in National SecurityReuters article - Mother sues AI chatbot company Character.AI, Google over son's suicideMedium article - A Small Step Towards Reproducing OpenAI o1: Progress Report on the Steiner Open Source ModelsThe Guardian article - Google's solution to accidental algorithmic racism: ban gorillasTIME article - Ethical AI Isn’t to Blame for Google’s Gemini DebacleLatacora article - The SOC2 Starting SevenGrandview Research market trends - Robotic Process Automation Market Trends
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
Winter is Coming for OpenAI
22 okt 2024· muckrAIkers
Brace yourselves, winter is coming for OpenAI - atleast, that's what we think. In this episode we look at OpenAI's recent massive funding round and ask "why would anyone want to fund a company that is set to lose net 5 billion USD for 2024?" We scrape through a whole lot of muck to find the meaningful signals in all this news, and there is a lot of it, so get ready!
(00:00) - Intro(00:28) - Hot off the press(02:43) - Why listen?(06:07) - Why might VCs invest?(15:52) - What are people saying(23:10) - How *is* OpenAI making money?(28:18) - Is AI hype dying?(41:08) - Why might big companies invest?(48:47) - Concrete impacts of AI(52:37) - Outcome 1: OpenAI as a commodity(01:04:02) - Outcome 2: AGI(01:04:42) - Outcome 3: best plausible case(01:07:53) - Outcome 1*: many ways to bust(01:10:51) - Outcome 4+: shock factor(01:12:51) - What's the muck(01:21:17) - Extended outro
Links
Reuters article - OpenAI closes $6.6 billion funding haul with investment from Microsoft and NvidiaGoldman Sachs report - GenAI: Too Much Spend, Too Little BenefitApricitas Economics article - The AI Investment BoomDiscussion of "The AI Investment Boom" on YCombinatorState of AI in 13 ChartsFortune article - OpenAI sees $5 billion loss in 2024 and soaring sales as big ChatGPT fee hikes planned, report says
More on AI Hype (Dying)
Latent Space article - The Winds of AI WinterArticle by Gary Marcus - The Great AI Retrenchment has BegunTimmermanReport article - AI: If Not Now, When? No, Really - When?MIT News article - Who Will Benefit from AI?Washington Post article - The AI Hype bubble is deflating. Now comes the hard part.Andreesen Horowitz article - Why AI Will Save the World
Other Sources
Human-Centered Artificial Intelligence Foundation Model Transparency IndexCointelegraph article - Europe gathers global experts to draft ‘Code of Practice’ for AIReuters article - Microsoft's VP of GenAI research to join OpenAITwitter post from Tim Brooks on joining DeepMindEdward Zitron article - The Man Who Killed Google Search
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
Open Source AI and 2024 Nobel Prizes
16 okt 2024· muckrAIkers
The Open Source AI Definition is out after years of drafting, will it reestablish brand meaning for the “Open Source” term? Also, the 2024 Nobel Prizes in Physics and Chemistry are heavily tied to AI; we scrutinize not only this year's prizes, but also Nobel Prizes as a concept.
(00:00) - Intro(00:30) - Hot off the press(03:45) - Open Source AI background(10:30) - Definitions and changes in RC1(18:36) - “Business source”(22:17) - Parallels with legislation(26:22) - Impacts of the OSAID(33:58) - 2024 Nobel Prize Context(37:21) - Chemistry prize(45:06) - Physics prize(50:29) - Takeaways(52:03) - What’s the real muck?(01:00:27) - Outro
Links
Open Source AI Definition, Release Candidate 1OSAID RC1 announcementAll Nobel Prizes 2024
More Reading on Open Source AI
Kairos.FM article - Open Source AI is a lie, but it doesn't have to beThe Register article - The open source AI civil war approachesMIT Technology Review article - We finally have a definition for open-source AI
On Nobel Prizes
Paper - Access to Opportunity in the Sciences: Evidence from the Nobel LaureatesPhysics prize - scientific background, popular infoChemistry prize - scientific background, popular infoReuters article - Google's Nobel prize winners stir debate over AI researchWikipedia article - Nobel disease
Other Sources
Pivot.ai article - People are ‘blatantly stealing my work,’ AI artist complainsPaper - GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language ModelsPaper - Reclaiming AI as a Theoretical Tool for Cognitive Science | Computational Brain & Behavior
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
SB1047
30 sep 2024· muckrAIkers
Why is Mark Ruffalo talking about SB1047, and what is it anyway? Tune in for our thoughts on the now vetoed California legislation that had Big Tech scared.
(00:00) - Intro(00:31) - Updates from a relatively slow week(03:32) - Disclaimer: SB1047 vetoed during recording (still worth a listen)(05:24) - What is SB1047(12:30) - Definitions(17:18) - Understanding the bill(28:42) - What are the players saying about it? (46:44) - Addressing critiques(55:59) - Open Source(01:02:36) - Takeaways(01:15:40) - Clarification on impact to big tech(01:18:51) - Outro
LinksSB1047 legislation pageSB1047 CalMatters pageNewsom vetoes SB1047CAIS newsletter on SB1047Prominent AI nerd letterAnthropic's letterSB1047 ~explainer
Additional SB1047 Related Coverage
Opposition to SB1047 'makes no sense'Newsom on SB1047Andreesen Horowitz on SB1047Classy move by DanEx-OpenAI employee says Altman doesn't want regulation
Other Sources
o1 doesn't measure up in new benchmark paperOpenAI losses and gainsOpenAI crypto hack"Murati out" -Mira Murati, probablyAltman pitching datacenters to White HouseSam Altman, 'podcast bro'Paper: Contract Design with Safety Inspections
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren
OpenAI's o1, aka. Strawberry
23 sep 2024· muckrAIkers
OpenAI's new model is out, and we are going to have to rake through a lot of muck to get the value out of this one!
⚠ Opt out of LinkedIn's GenAI scraping ➡️ https://lnkd.in/epziUeTi
(00:00) - Intro(00:25) - Other recent news(02:57) - Hot off the press(03:58) - Why might someone care?(04:52) - What is it?(06:49) - How is it being sold?(10:45) - How do they explain it, technically?(27:09) - Reflection AI Drama(40:19) - Why do we care?(46:39) - Scraping away the muck
Note: at around 32 minutes, Igor says the incorrect Llama model version for the story he is telling. Jacob dubbed over those mistakes with the correct versioning.
Links relating to o1
OpenAI blogpostSystem card webpageGitHub collection of o1 related mediaAMA Twitter threadFrancois Chollet Tweet on reasoning and o1The academic paper doing something very similar to o1
Other stuff we mention
OpenAI's huge valuation hinges on upending corporate structureMeta acknowledges it’s scraping all public posts for AI trainingWhite House announces new private sector voluntary commitments to combat image-based sexual abuseSam Altman wants you to be gratefulThe Zuck is done apologizingIAPS report on technical safety research at AI companiesLlama2 70B is "about as good" as GPT-4 at summarization tasks
- Luisteren Nogmaals beluisteren Doorgaan Wordt afgespeeld...
- Later beluisteren Later beluisteren

Afleveringen

DeepSeek Minisode

Understanding AI World Models w/ Chris Canal

NeurIPS 2024 Wrapped 🌯

OpenAI's o1 System Card, Literally Migraine Inducing

How to Safely Handle Your AGI

The End of Scaling?

US National Security Memorandum on AI, Oct 2024

Understanding Claude 3.5 Sonnet (New)

Winter is Coming for OpenAI

Open Source AI and 2024 Nobel Prizes

SB1047

OpenAI's o1, aka. Strawberry