Last Week in AI

Skynet Today

Weekly summaries and discussion about the most interesting developments in AI, deep learning, robotics, and more! read less
TechnologyTechnology

Episodes

#180 - Ideogram v2, Imagen 3, AI in 2030, Agent Q, SB 1047
6d ago
#180 - Ideogram v2, Imagen 3, AI in 2030, Agent Q, SB 1047
Our 180th episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) If you would like to get a sneak peek and help test Andrey's generative AI application, go to Astrocade.com to join the waitlist and the discord. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai Episode Highlights: Ideogram AI's new features, Google's Imagine 3, Dream Machine 1.5, and Runway's Gen3 Alpha Turbo model advancements.Perplexity's integration of Flux image generation models and code interpreter updates for enhanced search results. Exploration of the feasibility and investment needed for scaling advanced AI models like GPT-4 and Agent Q architecture enhancements.Analysis of California's AI regulation bill SB1047 and legal issues related to synthetic media, copyright, and online personhood credentials. Timestamps + Links: (00:00:00) Intro / Banter(00:01:08) Response to Listener Comments / CorrectionsTools & Apps (00:03:58) Ideogram AI expands its features with v2 model and color palette options(00:07:48) Google Releases Powerful AI Image Generator You Can Use for Free(00:11:41) Perplexity adds Flux.1 model for Pro users alongside Playground v3 update(00:13:58) Luma drops Dream Machine 1.5 — here’s what’s new(00:17:49) Runway’s Gen-3 Alpha Turbo is here and can make AI videos faster than you can type(00:20:21) Perplexity’s latest update improves code interpreter, charts included Applications & Business (00:24:14) AMD buying server maker ZT Systems for $4.9 billion as chipmakers strengthen AI capabilities(00:28:55) Ars Technica content is now available in OpenAI services(00:34:08) Anysphere, a GitHub Copilot rival, has raised $60M Series A at  $400M valuation from a16z, Thrive, sources say00:38:32 Stability AI appoints new Chief Technology Officer(00:41:45) Cruise’s robotaxis are coming to the Uber app in 2025 Projects & Open Source (00:44:16) AI21 Introduces the Jamba Model Family: The most powerful and efficient long-context models for the enterprise(00:53:47) Microsoft reveals Phi-3.5 — this new small AI model outperforms Gemini and GPT-4o(00:57:33) Nvidia’s Llama-3.1-Minitron 4B is a small language model that punches above its weight(01:00:58) Open source Dracarys models ignite generative AI fired coding Research & Advancements (01:12:35) Can AI Scaling Continue Through 2030?(01:15:35) Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents(01:23:58) Transformers to SSMs: Distilling Quadratic Knowledge to Subquadratic Models(01:31:18) Loss of plasticity in deep continual learning Policy & Safety (01:38:20) California weakens bill to prevent AI disasters before final vote, taking advice from Anthropic(01:48:14) Personhood credentials: Artificial intelligence and the value of privacy-preserving tools to distinguish who is real online(01:52:44) Showing SAE Latents Are Not Atomic Using Meta-SAEs Synthetic Media & Art (01:58:33) Authors sue Claude AI chatbot creator Anthropic for copyright infringement(01:59:32) Artists’ lawsuit against Stability AI and Midjourney gets more punch (02:01:43) Outro
#179 - Grok 2, Gemini Live, Flux, FalconMamba, AI Scientist
20-08-2024
#179 - Grok 2, Gemini Live, Flux, FalconMamba, AI Scientist
Our 179th episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) If you would like to get a sneak peek and help test Andrey's generative AI application, go to Astrocade.com to join the waitlist and the discord. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai Episode Highlights: - Grok 2's beta release features new image generation using Black Forest Labs' tech. - Google introduces Gemini Voice Chat Mode available to subscribers and integrates it into Pixel Buds Pro 2. - Huawei's Ascend 910C AI chip aims to rival NVIDIA's H100 amidst US export controls. - Overview of potential risks of unaligned AI models and skepticism around SingularityNet's AGI supercomputer claims. Timestamps + Links: (00:00:00) Intro / Banter(00:02:15) Response to listener comments / correctionsTools & Apps (00:04:24) Grok-2 is out in beta, now with added AI image generation(00:11:28) OpenAI reveals an updated GPT-4o model - but can't quite explain how it's better(00:13:48) Google Gemini’s voice chat mode is here(00:16:18) Google’s Pixel Buds Pro 2 bring Gemini to your ears(00:19:55) Google’s AI-generated search summaries change how they show their sources(00:23:13) Prompt Caching is Now Available on the Anthropic API for Specific Claude Models Applications & Business (00:26:56) Meet Black Forest Labs, the startup powering Elon Musk’s unhinged AI image generator(00:26:56) Huawei readies new AI chip to challenge Nvidia in China, WSJ reports(00:37:53) ASML and Imec Announce High-NA Lithography Breakthrough(00:43:07) Chinese startup WeRide gets nod to test robotaxis with passengers in California(00:45:49) Perplexity’s popularity surges as AI search start-up takes on Google(00:51:55) Lisa Su formally welcomes Silo AI team to AMD after completing $665 million acquisition Projects & Open Source (00:54:31) FalconMamba 7B Released: The World’s First Attention-Free AI Model with 5500GT Training Data and 7 Billion Parameters(00:59:25) OpenAI has introduced SWE-bench Verified to evaluate AI performance(01:04:21) Nous Research presents Hermes 3 (01:11:07) New supercomputing network could lead to AGI, scientists hope, with 1st node coming online within weeks Research & Advancements (01:14:40) The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery(01:30:24) Imagen 3(01:32:48) The Data Addition Dilemma(01:37:35) LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs Policy & Safety (01:40:55) MIT researchers release a repository of AI risks(01:44:14) Elon Musk addresses power issues at xAI supercomputer facility in Memphis(01:46:52) FCC Proposes New Rules on AI-Powered RobocallsGovernments Adjust Policies Amid Flood of AI Record Requests Synthetic Media & Art (01:48:21) SAG-AFTRA Strikes Groundbreaking AI Digital Voice Replica Pact With Startup Firm Narrativ(01:51:52) How ‘Deepfake Elon Musk’ Became the Internet’s Biggest Scammer (01:56:21) AI Song Outro
#178 - More Not-Acquihires, More OpenAI drama, More LLM Scaling Talk
16-08-2024
#178 - More Not-Acquihires, More OpenAI drama, More LLM Scaling Talk
Our 178th episode with a summary and discussion of last week's big AI news! NOTE: this is a re-upload with fixed audio, my bad on the last one! - Andrey With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) If you would like to get a sneak peek and help test Andrey's generative AI application, go to Astrocade.com to join the waitlist and the discord. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai In this episode: - Notable personnel movements and product updates, such as Character.ai leaders joining Google and new AI features in Reddit and Audible. - OpenAI's dramatic changes with co-founder exits, extended leaves, and new lawsuits from Elon Musk. - Rapid advancements in humanoid robotics exemplified by new models from companies like Figure in partnership with OpenAI, achieving amateur-level human performance in tasks like table tennis. - Research advancements such as Google's compute-efficient inference models and self-compressing neural networks, showcasing significant reductions in compute requirements while maintaining performance. Timestamps + Links: (00:00:00) Intro / Banter(00:03:14) Response to listener comments / correctionsApplications & Business(00:06:56) Google’s hiring of Character.AI’s founders is the latest sign that part of the AI startup world is starting to implode(00:15:12) Investors in Adept AI will be paid back after Amazon hires startup’s top talent(00:22:36) AI chip start-up Groq’s value rises to $2.8bn as it takes on Nvidia(00:29:22) OpenAI co-founder Schulman leaves for Anthropic, Brockman takes extended leave(00:36:18) Elon Musk files new lawsuit against OpenAI and Sam Altman(00:41:40) Figure’s new humanoid robot leverages OpenAI for natural speech conversations(00:47:01) ASML, Tokyo Electron dodge new US chip export rules, for now(00:53:10) OpenAI reportedly leads $60M round for webcam startup Opal Tools & Apps(00:55:40) OpenAI cuts GPT-4o prices, launches Structured Outputs amidst price war with Google(01:02:08) Apple Intelligence could get a $20 Plus version(01:04:05) Audible is testing an AI-powered search feature (01:05:53) Reddit to test AI-powered search result pages Research & Advancements(01:06:35) Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters(01:16:27) Achieving Human Level Competitive Robot Table Tennis(01:20:19) Self-Compressing Neural Networks(01:28:30) Let Me Speak Freely? A Study on the Impact of Format Restrictions on Performance of Large Language Models(01:32:43) Berkeley Humanoid: A Research Platform for Learning-based Control Policy & Safety(01:33:35) METR announces results of study on comparative capabilities of humans and agents(01:39:35) ‘The Godmother of AI’ says California’s well-intended AI bill will harm the U.S. ecosystem(01:49:13) Google Monopolized Search Through Illegal Deals, Judge Rules(01:54:56) Amazon faces UK merger probe over $4B Anthropic AI investment(01:55:44) GPT-4o System Card (02:03:09) Outro
#177 - Instagram AI Bots, Noam Shazeer -> Google, FLUX.1, SAM2
11-08-2024
#177 - Instagram AI Bots, Noam Shazeer -> Google, FLUX.1, SAM2
Our 177th episode with a summary and discussion of last week's big AI news! With guest co-host Jon Krohn from the super data science podcast (https://www.superdatascience.com/podcast)! If you'd like to listen to the interview with Andrey, check out https://www.superdatascience.com/podcast If you would like to get a sneak peek and help test Andrey's generative AI application, go to Astrocade.com to join the waitlist and the discord. In this episode, hosts Andrey Kurenkov and Jon Krohn dive into significant updates and discussions in the AI world, including Instagram's new AI features, Waymo's driverless cars rollout in San Francisco, and NVIDIA’s chip delays. They also review Meta's AI Studio, character.ai CEO Noam Shazir's return to Google, and Google's Gemini updates. Additional topics cover NVIDIA's hardware issues, advancements in humanoid robots, and new open-source AI tools like Open Devon. Policy discussions touch on the EU AI Act, the U.S. stance on open-source AI, and investigations into Google and Anthropic. The impact of misinformation via deepfakes, particularly one involving Elon Musk, is also highlighted, all emphasizing significant industry effects and regulatory implications. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai (00:00:00) AI Song / Intro Banter(00:05:32) Response to listener comments / correctionsTools & Apps (00:10:16) Apple Intelligence to Miss Initial Launch of Upcoming iOS 18 Overhaul(00:16:35) Instagram starts letting people create AI versions of themselvesLighting round (00:22:49) Runway just dropped image-to-video in Gen3(00:25:41) Midjourney drops surprise v6.1 update — now humans look more real than ever(00:28:07) AI-Powered Necklace Will Be Your Friend for $99(00:30:06) Microsoft is adding AI-powered summaries to Bing search results Applications & Business (00:31:44) Character.AI CEO Noam Shazeer returns to Google(00:39:41) Perplexity is cutting checks to publishers following plagiarism accusationsLighting round (00:43:30) Nvidia reportedly delays its next AI chip due to a design flaw(00:41:08) Neura shows off humanoid robot 4NE-1(00:46:0) Yes, there are more driverless Waymos in S.F. Here’s how busy they are(00:57:27) Canva acquires Leonardo.ai to boost its generative AI efforts Projects & Open Source (00:59:19) Black Forest Labs Open-Source FLUX.1: A 12 Billion Parameter Rectified Flow Transformer Capable of Generating Images from Text Descriptions(01:01:59) Google releases new ‘open’ AI models with a focus on safetyLighting round (01:05:09) Stability AI releases super-fast model for 3D asset image generation(01:09:29) OpenDevin: An Open Platform for AI Software Developers as Generalist Agents Research & Advancements (01:12:10) Meta AI Introduces Meta Segment Anything Model 2 (SAM 2): The First Unified Model for Segmenting Objects Across Images and Videos(01:19:20) MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware ExpertsLighting round (01:25:00) AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?(01:26:19) Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement(01:31:15) Stretching Each Dollar: Diffusion Training from Scratch on a Micro-Budget Policy & Safety (01:33:03) World's First-Ever AI Law Now Enforced in Europe, Targeting US Tech Giants(01:39:12) White House says no need to restrict ‘open-source’ artificial intelligence — at least for nowLighting round (01:41:12) With Smugglers and Front Companies, China Is Skirting American A.I. Bans(01:44:03) UK antitrust body probes Google’s ties with AI rival Anthropic(01:45:20) Elon Musk posts deepfake of Kamala Harris that violates X policy (01:50:10) AI Outro
#176 - BIG WEEK for OSS! SearchGPT, Lamma 3.1 405B, Mistral Large 2
03-08-2024
#176 - BIG WEEK for OSS! SearchGPT, Lamma 3.1 405B, Mistral Large 2
Our 176th episode with a summary and discussion of last week's big AI news! NOTE: apologies for this episode coming out about a week late, things got in the way of editing it... With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris)   Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai (00:00:00) Intro Song(00:00:34) Intro BanterTools & Apps(00:03:39) OpenAI announces SearchGPT, its AI-powered search engine(00:08:03) Google gives free Gemini users access to its faster, lighter 1.5 Flash AI model(00:09:10) X launches underwhelming Grok-powered ‘More About This Account’ feature(00:11:36) Kuaishou Launches Full Beta Testing for 'Kling AI' to Global Users, Elevates Model Capabilities(00:13:39) Adobe rolls out more generative AI features to Illustrator and Photoshop(00:14:25) Meta AI gets new ‘Imagine me’ selfie feature Projects & Open Source(00:15:19) Meta releases open-source AI model it says rivals OpenAI, Google tech(00:28:23) Mistral AI Unveils Mistral Large 2, Beats Llama 3.1 on Code and Math(00:34:00) Groq’s open-source Llama AI model tops leaderboard, outperforming GPT-4o and Claude in function calling(00:36:35) Apple shows off open AI prowess: new models outperform Mistral and Hugging Face offerings Applications & Business(00:40:25) Elon Musk wants Tesla to invest $5 billion into his newest startup, xAI — if shareholders approve(00:43:01) Nvidia said to be prepping Blackwell GPUs for Chinese market(00:46:28) Toronto AI company Cohere to indemnify customers who are sued for any copyright violations(00:49:09) AI startup Cohere raises US$500-million, valuing company at US$5.5-billion Research & Advancements(00:52:01) AI achieves silver-medal standard solving International Mathematical Olympiad problems(00:56:47) A Multimodal Automated Interpretability Agent(01:00:56) MINT-1T: Scaling Open-Source Multimodal Data by 10x: A Multimodal Dataset with One Trillion Tokens Policy & Safety(01:02:56) Improving Model Safety Behavior with Rule-Based Rewards(01:06:39) Senators demand OpenAI detail efforts to make its AI safe(01:10:59) OpenAI reassigns top AI safety executive Aleksandr Madry to role focused on AI reasoning(01:13:08) As new tech threatens jobs, Silicon Valley promotes no-strings cash aid(01:17:33) Democratic senators seek to reverse Supreme Court ruling that restricts federal agency power Synthetic Media & Art(01:20:58) Video game performers will go on strike over artificial intelligence concerns (01:23:03) Outro(01:23:58) AI Song
#175 - GPT-4o Mini, OpenAI's Strawberry, Mixture of A Million Experts
25-07-2024
#175 - GPT-4o Mini, OpenAI's Strawberry, Mixture of A Million Experts
Our 175th episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) In this episode of Last Week in AI, hosts Andrey Kurenkov and Jeremy Harris explore recent AI advancements including OpenAI's release of GPT 4.0 Mini and Mistral’s open-source models, covering their impacts on affordability and performance. They delve into enterprise tools for compliance, text-to-video models like Hyper 1.5, and YouTube Music enhancements. The conversation further addresses AI research topics such as the benefits of numerous small expert models, novel benchmarking techniques, and advanced AI reasoning. Policy issues including U.S. export controls on AI technology to China and internal controversies at OpenAI are also discussed, alongside Elon Musk's supercomputer ambitions and OpenAI’s Prover-Verify Games initiative.   Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai   Timestamps + links: (00:00:00) AI Song Intro(00:00:40) Intro / BanterTools & Apps(00:03:57) OpenAI unveils GPT-4o mini, a small AI model powering ChatGPT(00:11:38) Meet Haiper 1.5, the new AI video generation model challenging Sora, Runway(00:16:32) Anthropic releases Claude app for Android(00:18:59) Google Vids is available to test out Gemini AI-created video presentations(00:20:27) YouTube Music sound search rolling out, AI ‘conversational radio’ in testing  Applications & Business(00:23:30) OpenAI working on new reasoning technology under code name ‘Strawberry’(00:30:45) Inside Elon Musk’s Mad Dash To Build A Giant xAI Supercomputer In Memphis(00:37:15) Apple, NVIDIA and Anthropic reportedly used YouTube transcripts without permission to train AI models(00:41:05) After Tesla and OpenAI, Andrej Karpathy’s startup aims to apply AI assistants to education(00:43:40) Menlo Ventures and Anthropic team up on a $100M AI fund Projects & Open Source(00:46:27) Mistral releases Codestral Mamba for faster, longer code generation(00:50:36) Mistral AI and NVIDIA Unveil Mistral NeMo 12B, a Cutting-Edge Enterprise AI Model(00:52:51) Hugging Face Releases SmoLLM, a Series of Small Language Models, Beats Qwen2 and Phi 1.5(00:56:11) Stable Diffusion 3 License Revamped Amid Blowback, Promising Better Model Research & Advancements(01:01:49) FlashAttention-3 unleashes the power of H100 GPUs for LLMs(01:06:38) Mixture of A Million Experts(01:12:51) AutoBencher: Creating Salient, Novel, Difficult Datasets for Language Models(01:18:23) SpreadsheetLLM: Encoding Spreadsheets for Large Language >Models Policy & Safety(01:20:50) Prover-Verifier Games improve legibility of language model outputs(01:28:05) Trump allies draft AI order to launch ‘Manhattan Projects’ for defense(01:34:40) On scalable oversight with weak LLMs judging strong LLMs(01:36:24) Google, Microsoft offer Nvidia chips to Chinese companies, the Information reports(01:38:26) U.S. planning 'draconian' sanctions against China's semiconductor industry: Report(01:48:47) OpenAI illegally barred staff from airing safety risks, whistleblowers say (01:44:59) Outro + AI Song
#174 - Odyssey Text-to-Video, Groq LLM Engine, OpenAI Security Issues
17-07-2024
#174 - Odyssey Text-to-Video, Groq LLM Engine, OpenAI Security Issues
Our 174rd episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) In this episode of Last Week in AI, we delve into the latest advancements and challenges in the AI industry, highlighting new features from Figma and Quora, regulatory pressures on OpenAI, and significant investments in AI infrastructure. Key topics include AMD's acquisition of Silo AI, Elon Musk's GPU cluster plans for XAI, unique AI model training methods, and the nuances of AI copying and memory constraints. We discuss developments in AI's visual perception, real-time knowledge updates, and the need for transparency and regulation in AI content labeling and licensing. See full episode notes here. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai   Timestamps + links: (00:00:00) Intro AI Song(00:00:41) Pre News BanterTools & Apps(00:07:09) Odyssey Building 'Hollywood-Grade' AI Text-to-Video Model to Compete With Sora, Gen-3 Alpha(00:10:28) Anthropic’s Claude adds a prompt playground to quickly improve your AI apps(00:15:06) Figma pauses its new AI feature after Apple controversy(00:18:30) Quora’s Poe now lets users create and share web apps(00:20:54) Suno launches iPhone app — now you can make AI music on the go Applications & Business(00:21:42) Groq unveils lightning-fast LLM engine; developer base rockets past 280K in 4 months(00:27:03) Microsoft and Apple ditch OpenAI board seats amid regulatory scrutiny(00:29:39) OpenAI and Arianna Huffington are working together on an ‘AI health coach’(00:33:38) AI coding startup Magic seeks $1.5-billion valuation in new funding round, sources say(00:37:01) Sequoia and Andreessen Horowitz Clash Over AI Chip Supplies Amid Gen AI Boom(00:43:30) Elon Musk Reveals Plans To Make World’s “Most Powerful” 100,000 NVIDIA GPU AI Cluster(00:46:25) AMD plans to acquire Silo AI in $665 million deal(00:48:00) AI robotics startup raises US$300 million, including from Jeff Bezos(00:52:11) Intel begins groundwork on Magdeburg chip fab despite 13 remaining regulatory and environmental objections Research & Advancements(00:55:21) Learning to (Learn at Test Time): RNNs with Expressive Hidden States(01:03:12) Data curation via joint example selection further accelerates multimodal learning(01:09:11) CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation(01:13:25) Just read twice: closing the recall gap for recurrent language models(01:15:25) CodeUpdateArena: Benchmarking Knowledge Editing on API Updates(01:18:31) Composable Interventions for Language Models(01:24:09) Mind-reading AI recreates what you're looking at with amazing accuracy Policy & Safety(01:26:49) Covert Malicious Finetuning(01:31:23) OpenAI’s week of security issues(01:36:39) Here’s how OpenAI will determine how powerful its AI systems are(01:39:56) Me, Myself and AI: The Situational Awareness Dataset for LLMs(01:44:34) Exclusive: OpenAI partners with Los Alamos to study AI in the lab(01:47:36) Judge dismisses coders’ DMCA claims against Microsoft, OpenAI and GitHub(01:49:55) A former OpenAI safety employee said he quit because the company's leaders were 'building the Titanic' and wanted 'newer, shinier' things to sell Synthetic Media & Art(01:52:46) Vimeo joins YouTube and TikTok in launching new AI content labels(01:54:50) Tech Startup Aims to Help Media License Content for AI Training(01:57:23) Etsy adds AI-generated item guidelines in new seller policy (01:59:44) Bumble users can now report profiles that use AI-generated photos (02:02:05) Outro + AI Song
#173 - Gemini Pro, Llama 400B, Gen-3 Alpha, Moshi, Supreme Court
07-07-2024
#173 - Gemini Pro, Llama 400B, Gen-3 Alpha, Moshi, Supreme Court
Our 173rd episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) See full episode notes here. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai In this episode of Last Week in AI, we explore the latest advancements and debates in the AI field, including Google's release of Gemini 1.5, Meta's upcoming LLaMA 3, and Runway's Gen 3 Alpha video model. We discuss emerging AI features, legal disputes over data usage, and China's competition in AI. The conversation spans innovative research developments, cost considerations of AI architectures, and policy changes like the U.S. Supreme Court striking down Chevron deference. We also cover U.S. export controls on AI chips to China, workforce development in the semiconductor industry, and Bridgewater's new AI-driven financial fund, evaluating the broader financial and regulatory impacts of AI technologies.   Timestamps + links: (00:00:00) Intro / BanterTools & Apps(00:03:24) Google opens up Gemini 1.5 Flash, Pro with 2M tokens to the public(00:08:47) Meta is about to launch its biggest Llama model yet — here’s why it’s a big deal(00:12:38) Runway’s Gen-3 Alpha AI video model now available – but there’s a catch(00:16:28) This is Google AI, and it's coming to the Pixel 9(00:17:30) AI Firm ElevenLabs Sets Audio Reader Pact With Judy Garland, James Dean, Burt Reynolds and Laurence Olivier Estates(00:20:06) Perplexity’s ‘Pro Search’ AI upgrade makes it better at math and research(00:23:12) Gemini’s data-analyzing abilities aren’t as good as Google claims Applications & Business(00:26:38) Quora’s Chatbot Platform Poe Allows Users to Download Paywalled Articles on Demand(00:32:04) Huawei and Wuhan Xinxin to develop high-bandwidth memory chips amid US restrictions(00:34:57) Alibaba’s large language model tops global ranking of AI developer platform Hugging Face(00:39:01) Here comes a Meta Ray-Bans challenger with ChatGPT-4o and a camera(00:43:35) Apple’s Phil Schiller is reportedly joining OpenAI’s board(00:47:26) AI Video Startup Runway Looking to Raise $450 Million Projects & Open Source(00:48:10) Kyutai Open Sources Moshi: A Real-Time Native Multimodal Foundation AI Model that can Listen and Speak(00:50:44) MMEvalPro: Calibrating Multimodal Benchmarks Towards Trustworthy and Efficient Evaluation(00:53:47) Anthropic Pushes for Third-Party AI Model Evaluations(00:57:29) Mozilla Llamafile, Builders Projects Shine at AI Engineers World's Fair Research & Advancements(00:59:26) Researchers upend AI status quo by eliminating matrix multiplication in LLMs(01:05:55) AI Agents That Matter(01:12:09) WARP: On the Benefits of Weight Averaged Rewarded Policies(01:17:20) Scaling Synthetic Data Creation with 1,000,000,000 Personas(01:24:16) Found in the Middle: Calibrating Positional Attention Bias Improves Long Context Utilization Policy & Safety(01:26:32) With Chevron’s demise, AI regulation seems dead in the water(01:33:40) Nvidia to make $12bn from AI chips in China this year despite US controls(01:37:52) Uncle Sam relies on manual processes to oversee restrictions on Huawei, other Chinese tech players(01:40:57) U.S. government addresses critical workforce shortages for the semiconductor industry with new program(01:42:42) Bridgewater starts $2 billion fund that uses machine learning for decision-making and will include models from OpenAI, Anthropic and Perplexity (01:47:57) Outro
#172 - Claude and Gemini updates, Gemma 2, GPT-4 Critic
01-07-2024
#172 - Claude and Gemini updates, Gemma 2, GPT-4 Critic
Our 172nd episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ If you would like to become a sponsor for the newsletter, podcast, or both, please fill out this form. Email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai (00:00:00) Intro / BanterTools & Apps (00:03:02) Anthropic Debuts Collaboration Tools for Claude AI Assistant(00:08:32) Google rolls out Gemini side panels for Gmail and other Workspace apps(00:12:30) OpenAI delays rolling out its 'Voice Mode' to July(00:15:40) OpenAI’s ChatGPT for Mac is now available to all users(00:17:27) Waymo ditches the waitlist and opens up its robotaxis to everyone in San Francisco(00:18:53) Figma announces big redesign with AI Applications & Business (00:21:37) Meet Sohu: The World’s First Transformer Specialized Chip ASIC(00:29:42) Huawei Has Reportedly Invested Billions In An R&D Facility That Will Allow It To Develop Advanced Chipmaking Machinery Similar To ASML & Others(00:32:17) China's ByteDance working with Broadcom to develop advanced AI chip, sources say(00:35:35) Chinese AI firms woo OpenAI users as US company plans API restrictions(00:39:45) OpenAI walks back controversial stock sale policies, will treat current and former employees the same Projects & Open Source (00:43:42) Meta Large Language Model Compiler: Foundation Models of Compiler Optimization(00:47:54) Google’s Gemma 2 series launches with not one, but two lightweight model options—a 9B and 27B(00:48:50) ESM3: Simulating 500 million years of evolution with a language model Research & Advancements (00:56:57) Finding GPT-4’s mistakes with GPT-4(01:03:30) Chinese-built ChatGLM exceeds GPT-4 Across Several Benchmarks(01:07:15) Performances are plateauing, let's make the leaderboard steep again(01:11:18) Structural mechanism of bridge RNA-guided recombination(01:15:01) Reconciling Kaplan and Chinchilla Scaling Laws Policy & Safety (01:17:42) Safety Alignment Should Be Made More Than Just a Few Tokens Deep(01:23:02) Y Combinator rallies start-ups against California’s AI safety bill(01:28:20) Pro-Kigali propagandists caught using Artificial Intelligence tools(01:21:40) Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI(01:35:08) Adversaries Can Misuse Combinations of Safe ModelsMitigating Skeleton Key, a new type of generative AI jailbreak technique Synthetic Media & Art (01:39:35) Music labels sue AI music generators for copyright infringement(01:42:43) YouTube is trying to make AI music deals with major record labels(01:45:07) Toys ‘R’ Us Debuts First Video Ad Using Sora, OpenAI’s Text-to-Video Tool (01:49:12) Outro + AI Song
#171 - Apple Intelligence, Dream Machine, SSI Inc
24-06-2024
#171 - Apple Intelligence, Dream Machine, SSI Inc
Our 171st episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) Feel free to leave us feedback here. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + Links: (00:00:00) Intro / BanterTools & Apps(00:03:13) Apple Intelligence: every new AI feature coming to the iPhone and Mac(00:10:03) ‘We don’t need Sora anymore’: Luma’s new AI video generator Dream Machine slammed with traffic after debut(00:14:48) Runway unveils new hyper realistic AI video model Gen-3 Alpha, capable of 10-second-long clips(00:18:21) Leonardo AI image generator adds new video mode — here’s how it works(00:22:31) Anthropic just dropped Claude 3.5 Sonnet with better vision and a sense of humor Applications & Business(00:28:23 ) Sam Altman might reportedly turn OpenAI into a regular for-profit company(00:31:19) Ilya Sutskever, Daniel Gross, Daniel Levy launch Safe Superintelligence Inc.(00:38:53) OpenAI welcomes Sarah Friar (CFO) and Kevin Weil (CPO)(00:41:44) Report: OpenAI Doubled Annualized Revenue in 6 Months(00:44:30) AI startup Adept is in deal talks with Microsoft(00:48:55) Mistral closes €600m at €5.8bn valuation with new lead investor(00:53:12) Huawei Claims Ascend 910B AI Chip Manages To Surpass NVIDIA’s A100, A Crucial Alternative For China(00:56:58) Astrocade raises $12M for AI-based social gaming platform Projects & Open Source(01:01:03) Announcing the Open Release of Stable Diffusion 3 Medium, Our Most Sophisticated Image Generation Model to Date(01:05:53) Meta releases flurry of new AI models for audio, text and watermarking(01:09:39) ElevenLabs unveils open-source creator tool for adding sound effects to videos Research & Advancements(01:12:02) Samba: Simple Hybrid State Space Models for Efficient Unlimited Context Language Modeling(01:22:07) Improve Mathematical Reasoning in Language Models by Automated Process Supervision(01:28:01) Introducing Lamini Memory Tuning: 95% LLM Accuracy, 10x Fewer Hallucinations(01:30:32) An Empirical Study of Mamba-based Language Models(01:31:57) BERTs are Generative In-Context Learners(01:33:33) SELFGOAL: Your Language Agents Already Know How to Achieve High-level Goals Policy & Safety(01:35:16) Sycophancy to subterfuge: Investigating reward tampering in language models(01:42:26) Waymo issues software and mapping recall after robotaxi crashes into a telephone pole(01:45:53) Meta pauses AI models launch in Europe(01:46:44) Refusal in Language Models Is Mediated by a Single DirectionSycophancy to subterfuge: Investigating reward tampering in language models(01:51:38) Huawei exec concerned over China’s inability to obtain 3.5nm chips, bemoans lack of advanced chipmaking tools Synthetic Media & Art(01:55:07) It Looked Like a Reliable News Site. It Was an A.I. Chop Shop.(01:57:39) Adobe overhauls terms of service to say it won’t train AI on customers’ work(01:59:31) Buzzy AI Search Engine Perplexity Is Directly Ripping Off Content From News Outlets (02:02:23) Outro + AI Song
#170 - new Sora rival, OpenAI robotics, understanding GPT4, AGI by 2027?
09-06-2024
#170 - new Sora rival, OpenAI robotics, understanding GPT4, AGI by 2027?
Our 170th episode with a summary and discussion of last week's big AI news! With hosts Andrey Kurenkov (https://twitter.com/andrey_kurenkov) and Jeremie Harris (https://twitter.com/jeremiecharris) Feel free to leave us feedback here. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + Links: Tools & Apps(00:03:33) KLING is the latest AI video generator that could rival OpenAI's Sora(00:09:16) ‘Apple Intelligence’ will automatically choose between on-device and cloud-powered AI(00:12:21) Udio introduces new udio-130 music generation model and more advanced features(00:14:38) Perplexity AI’s new feature will turn your searches into shareable pages(00:16:35) ElevenLabs’ AI generator makes explosions or other sound effects with just a prompt(00:18:37) Google’s updated AI-powered NotebookLM expands to India, UK and over 200 other countries Applications & Business(00:19:40) OpenAI is restarting its robotics research group(00:25:01) Saudi fund invests in China effort to create rival to OpenAI(00:29:34) UAE seeks ‘marriage’ with US over artificial intelligence deals(00:33:01) Zoox to test self-driving cars in Austin and Miami (00:35:49) Microsoft Lays Off 1,500 Workers, Blames "AI Wave"(00:38:28) Avengers, assemble—Google, Intel, Microsoft, AMD and more team up to develop an interconnect standard to rival Nvidia's NVLink Projects & Open Source(00:40:39) GLM-4-9B-Chat-1M(00:46:37) Hugging Face and Pollen Robotics show off first project: an open source robot that does chores(00:49:40) Zyphra debuts Zyda, a 1.3T language modeling dataset it claims outperforms Pile, C4, arxiv(00:51:59) Stability AI debuts new Stable Audio Open for sound design Research & Advancements(00:54:05) Scaling and evaluating sparse autoencoders(01:04:54) Improving Alignment and Robustness with Short Circuiting(01:12:11) Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach(01:16:20) GPT-4 didn't ace the bar exam after all, MIT research suggests — it didn't even break the 70th percentile Policy & Safety(01:20:11) Former OpenAI researcher foresees AGI reality in 2027(01:28:03) OpenAI Insiders Warn of a ‘Reckless’ Race for Dominance(01:33:52) Testing and mitigating elections-related risks(01:36:26) Teams of LLM Agents can Exploit Zero-Day Vulnerabilities Synthetic Media & Art(01:43:23) The Uncanny Rise of the World's First AI Beauty Pageant (01:46:25) Outro + AI Song
#169 - Google's Search Errors, OpenAI news & DRAMA, new leaderboards
03-06-2024
#169 - Google's Search Errors, OpenAI news & DRAMA, new leaderboards
Our 168th episode with a summary and discussion of last week's big AI news! Feel free to leave us feedback here: https://forms.gle/ngXvXZpNJxaAprDv6 Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + Links: (00:00:00) Intro / Banter(00:02:55) Response to listener comments / correctionsTools & Apps (00:04:33) Google’s A.I. Search Errors Cause a Furor Online(00:10:56) Telegram gets an in-app Copilot bot(00:13:13) Opera is adding Google's Gemini AI to its browser(0016:13) Amazon plans to give Alexa an AI overhaul — and a monthly subscription price(00:19:15) Microsoft Edge will translate and dub YouTube videos as you’re watching them(00:21:12) Iyo thinks its gen AI earbuds can succeed where Humane and Rabbit stumbled Applications & Business (00:24:57) PwC agrees deal to become OpenAI's first reseller and largest enterprise user(00:30:07) Vox Media and The Atlantic sign content deals with OpenAI(00:36:27) OpenAI launches programs making ChatGPT cheaper for schools and nonprofits(00:40:03) Huawei patent reveals 3nm-class process technology plans — China continues to move forward despite US sanctions(00:44:32) Nvidia, Powered by A.I. Boom, Reports Soaring Revenue and Profits(00:48:16) Elon Musk’s xAI raises $6 billion in latest funding round Projects & Open Source (00:51:13) Scale AI publishes its first LLM Leaderboards, ranking AI model performance in specific domains(00:56:04) Cohere For AI Launches Aya 23, 8 and 35 Billion Parameter Open Weights Release(01:00:45) Who will make AlphaFold3 open source? Scientists race to crack AI model(01:04:07) Mistral releases Codestral, its first generative AI model for code Research & Advancements (01:09:23) The Road Less Scheduled(01:14:10) Training Compute of Frontier AI Models Grows by 4-5x per Year(01:21:33) gzip Predicts Data-dependent Scaling Laws(01:25:51) Neural Scaling Laws for Embodied AI(01:28:47) Contextual Position Encoding: Learning to Count What’s Important(01:33:09) New AI products much hyped but not much used, study says Policy & Safety (01:37:00) Ex-OpenAI board member reveals what led to Sam Altman's brief ousting(01:46:36) OpenAI researcher who resigned over safety concerns joins Anthropic(01:49:16) Leaked OpenAI Documents Show Sam Altman Was Clearly Aware of Silencing Former Employees(01:54:33) OpenAI Board Forms Safety and Security Committee(01:58:07) Robocaller Who Used AI to Clone Biden’s Voice Fined $6 Million(01:59:08) Hacker Releases Jailbroken "Godmode" Version of ChatGPT(02:00:46) China Creates $47.5 Billion Chip Fund to Back Nation’s Firms Synthetic Media & Art (02:02:23) Alphabet, Meta Offer Millions to Partner With Hollywood on AI (02:04:21) Outro + AI Song
#168 - OpenAI vs Scar Jo + safety researchers, MS AI updates, cool Anthropic research
28-05-2024
#168 - OpenAI vs Scar Jo + safety researchers, MS AI updates, cool Anthropic research
Our 168th episode with a summary and discussion of last week's big AI news! With guest host Gavin Purcell from AI for Humans podcast! Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + Links: (00:00:00) Intro / Banter + Response to listener comments / correctionsTools & Apps (00:08:00) OpenAI says Sky voice in ChatGPT will be paused after concerns it sounds too much like Scarlett Johansson(00:16:14) Microsoft’s Copilot assistant is getting a GPT-4o upgrade + Recall is Microsoft’s key to unlocking the future of PCs(00:21:36) ElevenLabs Launches AI-Voiced Screen Reader App(00:22:40) Adobe Lightroom gets a magic eraser, and it’s impressive(00:25:07) Microsoft, Khan Academy provide free AI assistant for all educators in US(00:27:40) Microsoft Paint is getting an AI-powered image generator that responds to your text prompts and doodles Applications & Business (00:29:16) OpenAI founders Sam Altman and Greg Brockman go on the defensive after top safety researchers quit(00:36:58) OpenAI, WSJ Owner News Corp Strike Content Deal Valued at Over $250 Million(00:41:27) CoreWeave Raises $7.5 Billion in Debt for AI Computing Push(00:44:13) Google announced Trillium, its sixth generation of Tensor processors.(00:45:09) Inflection AI reveals new team and plan to embed emotional AI in business bots(00:47:01) Data-labeling startup Scale AI raises $1B as valuation doubles to $13.8B Projects & Open Source (00:48:35) Abacus AI Releases Smaug-Llama-3-70B-Instruct: The New Benchmark in Open-Source Conversational AI Rivaling GPT-4 Turbo(00:52:24) Introducing New Chatbot Arena Category: Hard Prompts(00:54:56) Microsoft brings out a small language model that can look at pictures Research & Advancements (00:56:05) New Anthropic Research Sheds Light on AI's 'Black Box'(01:04:03) Chameleon: Mixed-Modal Early-Fusion Foundation Models(01:08:14) SpeechVerse: A Large-scale Generalizable Audio Language Model(01:09:05) CAT3D: Create Anything in 3D with Multi-View Diffusion Models(01:11:17) Coin3D: Controllable and Interactive 3D Assets Generation with Proxy-Guided Conditioning(01:12:10) SpeechGuard: Exploring the Adversarial Robustness of Multimodal Large Language Models Policy & Safety (01:15:01) World’s first major law for artificial intelligence gets final EU green light(01:17:18) Colorado governor signs sweeping AI regulation bill(01:22:10) Senators Propose $32 Billion in Annual A.I. Spending but Defer Regulation(01:23:25) Google DeepMind launches new framework to assess the dangers of AI models(01:25:05) Tech giants pledge AI safety commitments — including a ‘kill switch’ if they can’t mitigate risks Synthetic Media & Art (01:28:32) Sony Music warns tech companies over ‘unauthorized’ use of its content to train AI(01:32:34) Hollywood agency CAA aims to help stars manage their own AI likenesses(01:38:28) What Do You Do When A.I. Takes Your Voice? (01:42:01) Outro + AI Song
#167 - GPT-4o, Project Astra, Veo, OpenAI Departures, Interview with Andrey
19-05-2024
#167 - GPT-4o, Project Astra, Veo, OpenAI Departures, Interview with Andrey
Our 167th episode with a summary and discussion of last week's big AI news! With guest host Daliana Liu (https://www.linkedin.com/in/dalianaliu/) from The Data Scientist Show! And a special one-time interview with Andrey in the latter part of the podcast. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + links: Intro / BanterTools & Apps (00:03:42) OpenAI releases GPT-4o, a faster model that’s free for all ChatGPT users(00:12:06) Project Astra is the future of AI at Google(00:18:06) Google is redesigning its search engine — and it’s AI all the way down(00:19:39) Google unveils Veo and Imagen 3, its latest AI media creation models(00:23:36) Google Unveils Music AI Sandbox Making Loops From Prompts(00:26:27) Anthropic AI Launches a Prompt Engineering Tool that Generates Production-Ready Prompts in the Anthropic Console Applications & Business (00:31:02) OpenAI’s Chief Scientist and Co-Founder Is Leaving the Company(00:35:15) Mike Krieger joins Anthropic as Chief Product Officer(00:36:28) $16k G1 humanoid rises up to smash nuts, twist and twirl(00:41:02) GM's Cruise to start testing robotaxis in Phoenix area with human safety drivers on board(00:42:52) US agency probes Amazon-owned Zoox self-driving vehicles after two crashes(00:43:58) Waymo’s robotaxis under investigation after crashes and traffic mishaps Projects & Open Source (00:44:48) Introducing PaliGemma, Gemma 2, and an Upgraded Responsible AI Toolkit(00:46:24) Falcon 2: UAE’s Technology Innovation Institute Releases New AI Model Series, Outperforming Meta’s New Llama 3(00:48:00) License to Call: Introducing Transformers Agents 2.0 Research & Advancements (00:49:22) The Platonic Representation Hypothesis(00:53:08) SUTRA: Scalable Multilingual Language Model Architecture Policy & Safety (00:54:46) Bipartisan Senate bill on AI security would bolster voluntary cyber reporting processes(00:56:17) U.K. agency releases tools to test AI model safety(00:57:25) Protesters Are Fighting to Stop AI, but They’re Split on How to Do It Synthetic Media & Art (00:58:54) Google’s invisible AI watermark will help identify generative text and video(01:00:50) How One Author Pushed the Limits of AI Copyright(01:03:27) Stellaris gets an DLC about AI that features AI-created voices, director insists it's 'ethical' and 'we're pretty good at exploring dystopian sci-fi and don't want to end up there ourselves'(01:04:46) At the AI Film Festival, humanity triumphed over tech (01:06:37) Daliana Interviews Andrey(01:42:00) AI Outro Song
#166 - new AI song generator, Microsoft's GPT4 efforts, AlphaFold3, xLSTM, OpenAI Model Spec
12-05-2024
#166 - new AI song generator, Microsoft's GPT4 efforts, AlphaFold3, xLSTM, OpenAI Model Spec
Our 166th episode with a summary and discussion of last week's big AI news! Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + links: (00:00:00) Intro / BanterTools & Apps(00:04:23) ElevenLabs previews music-generating AI model(00:09:31) Mysterious “gpt2-chatbot” AI model appears suddenly, confuses experts(00:13:00) SoundHound AI and Perplexity Partner to Bring Online LLMs to Next Gen Voice Assistants Across Cars and IoT Devices(00:14:50) Stability AI sows gen AI discord with Stable Artisan(00:16:35) Apple Will Revamp Siri to Catch Up to Its Chatbot Competitors(00:18:54) Alibaba rolls out latest version of its large language model to meet robust AI demand Applications & Business(00:19:34) OpenAI and Stack Overflow partner to bring more technical knowledge into ChatGPT(00:17:31) New Microsoft AI model may challenge GPT-4 and Google Gemini(00:31:08) Wayve, an A.I. Start-Up for Autonomous Driving, Raises $1 Billion(00:32:00) Motional delays commercial robotaxi plans amid restructuring(00:33:54) The rise of the Chinese AI unicorns doing battle with OpenAI Projects & Open Source(00:35:25) Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models(00:40:12) DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model(00:44:31) OpenVoice V2: Evolving Multilingual Voice Cloning with Enhanced Style Control and Cross-Lingual Capabilities(00:45:20) Granite Code Models: A Family of Open Foundation Models for Code Intelligence(00:46:00) Hugging Face launches LeRobot open source robotics code library(00:48:50) Vibe-Eval: A new open and hard evaluation suite for measuring progress of multimodal language models Research & Advancements(00:50:02) Google DeepMind’s Groundbreaking AI for Protein Structure Can Now Model DNA(00:57:20) xLSTM: Extended Long Short-Term Memory(01:06:35) StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation(01:07:55) Chain of Thought Empowers Transformers to Solve Inherently Serial Problems(01:11:48) KAN: Kolmogorov-Arnold Networks Policy & Safety(01:13:20) US lawmakers unveil bill to make it easier to restrict exports of AI models(01:17:30) OpenAI’s Model Spec outlines some basic rules for AI(01:20:18) Robot dogs armed with AI-targeting rifles undergo US Marines Special Ops evaluation(01:25:15) OpenAI Releases ‘Deepfake’ Detector to Disinformation Researchers Synthetic Media & Art(01:28:15) Audible’s Test of AI-Voiced Audiobooks Tops 40,000 Titles(01:32:30) TikTok will automatically label AI-generated content created on platforms like DALL·E 3(01:33:23) Katy Perry's Fan-Made AI Image Is So Real It Fooled the World Into Thinking She Was at the Met Gala(01:35:32) South Korean woman falls for deepfake Elon Musk, loses $50K in romance scam (01:37:18) Why young Russian women appear so eager to marry Chinese men (01:40:18) AI Outro Song
#165 - Sora challenger, Astribot's S1, Med-Gemini, Refusal in LLMs
05-05-2024
#165 - Sora challenger, Astribot's S1, Med-Gemini, Refusal in LLMs
Our 165th episode with a summary and discussion of last week's big AI news! Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + links: Tools & Apps(00:01:27) GitHub releases an AI-powered tool aiming for a 'radically new way of building software'(00:07:05) China unveils Sora challenger able to produce videos from text similar to OpenAI tool, though much shorter(00:12:23) ChatGPT’s AI ‘memory’ can remember the preferences of paying customers(00:14:21) Rabbit R1 review: Avoid this AI gadget(00:18:30) Amazon Q, a generative AI-powered assistant for businesses and developers, is now generally available(00:19:54) Yelp’s Assistant AI bot will do all the talking to help users find service providers Applications & Business(00:21:31) Video of super-fast, super-smooth humanoid robot will drop your jaw(00:25:22) Tesla’s 2 million car Autopilot recall is now under federal scrutiny(00:29:32) Tesla shares soar as Elon Musk returns from China with FSD 'Game Changer'(00:32:11) OpenAI inks strategic tie-up with UK’s Financial Times, including content use(00:35:21) OpenAI Startup Fund quietly raises $15M(00:37:00) Huawei backs HBM memory manufacturing in China to sidestep crippling US sanctions that restrict AI development Research & Advancements(00:39:20) Capabilities of Gemini Models in Medicine(00:45:34) Let's Think Dot by Dot: Hidden Computation in Transformer Language Models(00:52:20) NExT: Teaching Large Language Models to Reason about Code Execution(00:55:08) SenseNova 5.0: China’s latest AI model surpasses OpenAI’s GPT-4(00:57:20) Octopus v4: Graph of language models(01:00:28) Better & Faster Large Language Models via Multi-token Prediction Policy & Safety(01:03:15) Refusal in LLMs is mediated by a single direction(01:09:19) Rishi Sunak promised to make AI safe. Big Tech’s not playing ball.(01:15:09) DOE Announces New Actions to Enhance America’s Global Leadership in Artificial Intelligence(01:18:21) The Chips Act is rebuilding US semiconductor manufacturing, so far resulting in $327 billion in announced projects(01:20:50) Analysis-Second global AI safety summit faces tough questions, lower turnout(01:24:03) Sam Altman, Jensen Huang, and more join the federal AI safety board Synthetic Media & Art(01:26:30) Air Head creators say OpenAI's Sora finicky to work with, needs hundreds of prompts, serious VFX work for under 2 minutes of cohesive story ↺(01:29:50) Eight newspaper publishers sue OpenAI over copyright infringement
#164 - Meta AI, Phi-3, OpenELM, Bollywood Deepfakes
30-04-2024
#164 - Meta AI, Phi-3, OpenELM, Bollywood Deepfakes
Our 164th episode with a summary and discussion of last week's big AI news! Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + links: Tools & Apps (00:04:02) Meta, in Its Biggest A.I. Push, Places Smart Assistants Across Its Apps(00:07:26) Microsoft launches Phi-3, its smallest AI model yet(00:15:35) The Ray-Ban Meta Smart Glasses have multimodal AI now(00:17:32) OpenAI winds down AI image generator that blew minds and forged friendships in 2022(00:18:44) Baidu claims 200 million users for Ernie chatbot after only 13 months(00:21:13) The new Adobe Photoshop gets an in-app image generator, major Generative Fill upgrades Applications & Business (00:22:22) Intel & The Pentagon Deepen Ties To Develop World’s Most Advanced Chips(00:27:58) Meta Says It Plans to Spend Billions More on A.I.(00:31:36) OpenAI CEO Sam Altman invests in solar power firm Exowatt to fuel AI datacenters(00:33:58) Google consolidates AI-focused DeepMind, Research teams(00:36:22) Microsoft and OpenAI bet $100 billion to free themselves from the shackles and overreliance on the world's most profitable semiconductor chip brand for AI chips Projects & Open Source (00:39:03) Apple releases OpenELM: small, open source AI models designed to run on-device(00:44:12) Snowflake launches Arctic, an open ‘mixture-of-experts’ LLM to take on DBRX, Llama 3 Research & Advancements (00:48:08) The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions(00:55:11) Groq’s breakthrough AI chip achieves blistering 800 tokens per second on Meta’s LLaMA 3(00:59:52) Microsoft shows off VASA-1, an AI framework that makes human headshots talk, sing(01:01:59) Intel Builds World’s Largest Neuromorphic System to Enable More Sustainable AI Policy & Safety (01:05:11) Deepfakes of Bollywood stars spark worries of AI meddling in India election(01:08:51) LLM Agents can Autonomously Exploit One-day Vulnerabilities(01:15:27) The Necessity of AI Audit Standards Boards(01:19:45) A Mechanism-Based Approach to Mitigating Harms from Persuasive Generative AI(01:22:45) COERCING LLMS TO DO AND REVEAL (ALMOST) ANYTHING(01:26:40) China acquired recently banned Nvidia chips in Super Micro, Dell servers, tenders show Synthetic Media & Art (01:29:08) Drake threatened with lawsuit over diss track featuring AI Tupac
#163 - Llama 3, Grok-1.5 Vision, new Atlas robot, RHO-1, Medium ban
24-04-2024
#163 - Llama 3, Grok-1.5 Vision, new Atlas robot, RHO-1, Medium ban
Our 163rd episode with a summary and discussion of last week's big AI news! Note: apology for this one coming out a few days late, got delayed in editing it -Andrey Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + links: Intro / BanterTools & Apps (00:02:16) Meta releases Llama 3, claims it’s among the best open models available(00:14:01) Elon Musk’s xAI Unveils Grok-1.5 Vision, Beats OpenAI’s GPT-4V(00:17:55) Reka releases Reka Core, its multimodal language model to rival GPT-4 and Claude 3 Opus(00:21:50) Cohere Compass Private Beta: A New Multi-Aspect Embedding Model(00:23:48) Amazon Music’s Maestro lets listeners make AI playlists(00:24:36) Snap plans to add watermarks to images created with its AI-powered tools Applications & Business (00:25:52) Boston Dynamics unveils new Atlas robot for commercial use(00:30:32) TSMC’s $65 billion bet still leaves US missing piece of chip puzzle(00:36:30) U.S. blacklists Intel's and Nvidia's key partner in China — three other Chinese firms also included in the blacklist for helping the military(00:38:37) Elon Musk says the next-generation Grok 3 model will require 100,000 Nvidia H100 GPUs to train(00:40:22) Dr. Andrew Ng appointed to Amazon’s Board of Directors(00:41:55) Collaborative Robotics Locks Up $100M, Latest Robot Startup To Raise Big Projects & Open Source (00:44:08) OpenEQA: Embodied Question Answering in the Era of Foundation Models(00:50:03) Introducing Idefics2: A Powerful 8B Vision-Language Model for the community Research & Advancements (00:51:21) RHO-1: Not All Tokens Are What You Need(00:57:21) Scaling Laws for Fine-Grained Mixture of Experts(01:03:20) Chinchilla Scaling: A replication attempt(01:07:18) China develops new light-based chiplet that could power artificial general intelligence — where AI is smarter than humans(01:10:45) OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments Policy & Safety (01:13:44) U.S. Commerce Secretary Gina Raimondo Announces Expansion of U.S. AI Safety Institute Leadership Team(01:17:18) NSA Publishes Guidance for Strengthening AI System Security(01:19:19) Foundational Challenges in Assuring Alignment and Safety of Large Language Models(01:24:11) Former OpenAI Board Member Calls for Audits of Top AI Companies(01:27:35) SoA survey reveals a third of translators and quarter of illustrators losing work to AI Synthetic Media & Art (01:30:25) Medium bans AI-generated content from its paid Partner Program
#162 - Udio Song AI, TPU v5, Mixtral 8x22, Mixture-of-Depths, Musicians sign open letter
15-04-2024
#162 - Udio Song AI, TPU v5, Mixtral 8x22, Mixture-of-Depths, Musicians sign open letter
Our 162nd episode with a summary and discussion of last week's big AI news! Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Timestamps + links: Tools & Apps (00:02:50) AI-Music Arms Race: Meet Udio, the Other ChatGPT for Music(00:07:42) Anthropic launches external tool use for Claude AI, enabling stock ticker integrations and more(00:11:51) Building LLMs for Code Repair(00:14:16) Early Reviews of Humane AI Pin Aren’t Impressed(00:16:23) Microsoft 365’s Copilot gets a GPT-4 Turbo upgrade and improved image generation(00:18:41) AI editing tools are coming to all Google Photos users Applications & Business (00:19:21) Google announces the Cloud TPU v5p, its most powerful AI accelerator yet(00:23:32) Meta unveils its newest custom AI chip as it races to catch up(00:27:27) Intel Unveils New AI Accelerator in Bid to Challenge Nvidia(00:30:46) Adobe Is Buying Videos for $3 Per Minute to Build AI Model(00:32:55) OpenAI transcribed over a million hours of YouTube videos to train GPT-4(00:36:23) Waymo will launch paid robotaxi service in Los Angeles on Wednesday(00:37:23) OpenAI removes Sam Altman's ownership of its Startup Fund Projects & Open Source (00:39:51) Mistral AI Stuns With Surprise Launch of New Mixtral 8x22B Model(00:43:54) Google updates its Gemma AI model family with variants for coding and research(00:47:04) Aurora-M: The First Open Source Multilingual Language Model Red-teamed according to the U.S. Executive Order Research & Advancements (00:52:08) Mixture-of-Depths: Dynamically allocating compute in transformer-based language models(00:57:41) Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention(01:03:31) Octopus v2: On-device language model for super agent(01:07:54) Bigger is not Always Better: Scaling Properties of Latent Diffusion Models(01:09:54) Many-shot Jailbreaking Policy & Safety (01:15:08) Schiff unveils AI training transparency measure(01:20:25) Linwei Ding was a Google software engineer. He was also a prolific thief of trade secrets, say prosecutors.(01:26:11) Responsible Reporting for Frontier AI Development(01:30:08) US govt wants to talk to tech companies about AI electricity demands — eyes nuclear fusion and fission(01:32:39) Washington state judge blocks use of AI-enhanced video as evidence in possible first-of-its-kind ruling(01:36:45) Trudeau announces $2.4 billion for AI-related investments Synthetic Media & Art (01:39:26) Billie Eilish, Pearl Jam, Nicki Minaj Among 200 Artists Calling for Responsible AI Music Practices Fun! (01:41:52) OpenAI's Sora just made its first music video and it's like a psychedelic trip
#161 - Claude 3 beats GPT-4, Stability CEO resigns, DBRX, TacticAI, UN resolution on AI
31-03-2024
#161 - Claude 3 beats GPT-4, Stability CEO resigns, DBRX, TacticAI, UN resolution on AI
Our 161st episode with a summary and discussion of last week's big AI news! Check out our sponsor, the SuperDataScience podcast. You can listen to SDS across all major podcasting platforms (e.g., Spotify, Apple Podcasts, Google Podcasts) plus there’s a video version on YouTube. Read out our text newsletter and comment on the podcast at https://lastweekin.ai/ Email us your questions and feedback at contact@lastweekin.ai and/or hello@gladstone.ai Note - one extra story we didn't get to but worth knowing from this week: ‘Totally surreal’: OpenAI shares first short films created with new AI tool Sora Timestamps + links: (00:00:00) Intro / BanterTools & Apps (00:05:20) Google starts testing AI overviews from SGE in main Google search interface(00:10:00) Adobe’s new GenStudio platform is an AI factory for advertisers(00:13:53) Claude-3 Haiku has reached GPT-4 level by our user preference(00:15:26) Microsoft Teams is getting smarter Copilot AI features(00:17:16) Samsung is beating Apple in the race to bring AI to smartphones(00:19:18) Elon Musk says all premium subscribers on X will gain access to AI chatbot Grok this week Applications & Business (00:22:36) Stability AI CEO resigns to ‘pursue decentralized AI’(00:26:43) Amazon spends $2.75 billion on AI startup Anthropic in its largest venture investment yet(00:31:16) Intel Gaudi 2 Accelerators Showcase Competitive Performance Per Dollar Against NVIDIA H100 In MLPerf 4.0 GenAI Benchmarks(00:33:43) Chip Startup Celestial AI Lands Massive $175M Series C(00:36:10) Accenture Invests in Sanctuary AI to Bring AI-Powered, Humanoid Robotics to Work Alongside Humans(00:37:44) Humanoid robots are joining the Mercedes-Benz workforce Projects & Open Source (00:38:40) Introducing DBRX: A New State-of-the-Art Open LLM(00:44:06) Common Corpus: A Large Public Domain Dataset for Training LLMs(00:46:07) DROID: A Large-Scale In-the-Wild Robot Manipulation Dataset(00:48:15) InternLM2 Technical Report Research & Advancements (00:51:49) TacticAI: an AI assistant for football tactics(00:55:44) On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trial(00:59:26) AutoDev: Automated AI-Driven Development(01:02:48) Reverse Training to Nurse the Reversal Curse(01:07:43) Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference(01:09:48) Gamba: Marry Gaussian Splatting with Mamba for single view 3D reconstruction Policy & Safety (01:11:12) General Assembly adopts landmark resolution on artificial intelligence(01:16:24) Israel Deploys Expansive Facial Recognition Program in Gaza(01:19:49) LINEARITY OF RELATION DECODING IN TRANSFORMER LANGUAGE MODELS(01:24:52) US Weighs Sanctioning Huawei’s Secretive Chinese Chip Network(01:27:51) New York City welcomes robotaxis — but only with safety drivers(01:28:40) The White House Puts New Guardrails on Government Use of AI Synthetic Media & Art (01:31:17) BBC Will Stop Using AI For ‘Doctor Who’ Promotion After Receiving Complaints