Openai
entities · 159 notes linked
Related: Large Language Models · Chatgpt · AI Safety · Google · Anthropic · Generative AI · AI Agents · Prompt Engineering
Notes
- $450 and 19 hours is all it takes to rival OpenAI's o1-preview — UC Berkeley Sky-T1-32B open-source reasoning model built for $450 in 19 hours
- 'If journalism is going up in smoke, I might as well get high off the fumes': confessions of a chatbot helper — Human annotators writing gold-standard training data for LLMs
- 'The illusion of thinking': Apple research finds AI models collapse and give up with hard puzzles — Apple study showing large reasoning models collapse on hard logic puzzles
- 10 things I learned from burning myself out with AI coding agents — Practical limits of AI coding agents after 50 real projects
- 255 — Paul Christiano's defense of RLHF research positive net impact
- 35 Best Resources To Learn Machine Learning — Curated list of free interactive ML and deep learning visualization tools
- 8 Google Employees Invented Modern AI. Here's the Inside Story — Origin story of "Attention Is All You Need" transformer paper at Google
- A Coder Considers the Waning Days of the Craft — Programmer's first-person reflection on GPT-4 displacing coding craft
- A Google researcher – who said she was fired after pointing out biases in AI – says companies won't 'self-regulate' because of the AI 'gold rush' — Timnit Gebru warns AI gold rush prevents corporate self-regulation on bias
- A Note From Ray Kurzweil on the Recent Call to Pause Work on AI More Powerful Than GPT-4 — Ray Kurzweil opposing FLI's AI pause letter citing vagueness and coordination failure
- A Very Short Introduction to Diffusion Models — Forward noise addition and learned reverse denoising for image generation
- A Wave Of Billion-Dollar Language AI Startups Is Coming — 2022 landscape survey of language AI startup ecosystem categories
- A brief history of diffusion, the tech at the heart of modern image-generating AI | TechCrunch — Origins and capabilities of diffusion models powering modern image generation AI
- AI #171: False Flag — Weekly roundup; OpenAI-linked PAC false-flag scandal
- AI #172: The First Fable — Weekly roundup excluding the Fable model itself
- AI #174: You're It — Weekly AI roundup, Claude Tag, medical scanners, agent security
- AI #175: The Fable Continues — Weekly AI roundup covering Fable's return aftermath
- AI 2027 — HN debate on AI 2027 scenario: AGI timelines, LLM limits, alignment risks
- AI 50 2025: AI Agents Move Beyond Chat — Forbes AI 50 marks shift from chatbots to full workflow automation
- AI Is a Waste of Time — AI tools as entertainment and time-wasting before productivity gains materialize
- AI can now model and design the genetic code for all domains of life with Evo 2 | Arc Institute — Largest open-source AI biology model trained on 9.3 trillion nucleotides
- AI hype is built on high test scores. Those tests are flawed. — LLM benchmark scores are brittle, anthropomorphized, and often measure memorization not capability
- AI is already linked to layoffs in the industry that created it | CNN Business — AI driving tech sector layoffs and workforce skill reshuffling
- AI is going to eliminate way more jobs than anyone realizes — Generative AI's massive disruptive impact on global labor markets and productivity
- Amazon warns employees not to share confidential information with ChatGPT after seeing cases where its answer 'closely matches existing material' from inside the company — Amazon restricts employee ChatGPT use over corporate data leakage risk
- An Opinionated Guide to ML Research — Practical advice on problem selection and research habits for ML researchers
- Artificial General Intelligence Is Already Here | NOEMA — Argument that current frontier LLMs have already achieved meaningful artificial general intelligence
- Artists Can Fight Back Against AI by Killing Art Generators From the Inside — Nightshade tool poisons AI art model training data through imperceptible pixel manipulation
- At TED AI 2023, experts debate whether we've created "the new electricity — TED AI 2023 first AI-only TED conference debates AGI benefits and risks
- AudioPen is a great web app for converting your voice into text notes | TechCrunch — Voice-to-text note-taking web app powered by OpenAI Whisper
- Can AI really be protected from text-based attacks? | TechCrunch — LLM prompt injection attacks are low-barrier and currently unpreventable
- Chat with your data using OpenAI, Pinecone, Airbyte and Langchain — RAG pipeline tutorial with Airbyte and Pinecone
- ChatGPT Gets Its "Wolfram Superpowers"! — ChatGPT plugin connecting to Wolfram Alpha and Language
- ChatGPT Makes OK Clinical Decisions—Usually — Study finds ChatGPT 72% accurate across clinical decision tasks using fictional vignettes
- ChatGPT and generative AI are booming, but the costs can be extraordinary — Compute economics of training and serving large language models
- ChatGPT is 'not particularly innovative,' and 'nothing revolutionary', says Meta's chief AI scientist — Yann LeCun argues ChatGPT is solid engineering not scientific breakthrough
- ChatGPT is not all you need. A State of the Art Review of large Generative AI models — Taxonomy of 2022-2023 generative AI models across all major modalities
- ChatGPT-4 Receives 'B' on Scott Aaronson's Quantum Information Science Final — GPT-4 scores B on honors quantum information science exam
- Chegg's stock plunges on fears of competition from ChatGPT — ChatGPT disrupts edtech as Chegg stock crashes 48% in one day
- Claude Fable 5 and Mythos 5: Capabilities — Fable 5 capability review, benchmarks, classifiers, reception
- Claude Sonnet 5 Is Not Frontier But Has Its Uses — Sonnet 5 system card review, cheaper faster non-frontier model
- Claude's Chrome plugin is now available to all paid users - Engadget — Anthropic's Claude Chrome plugin expands beyond Max plan subscribers
- Cost calculations for LLM providers — Comparative USD-per-million-token pricing table with LMSYS ELO scores
- DS-STAR: A state-of-the-art versatile data science agent — Iterative plan-verify data science agent handling heterogeneous file formats
- Deep Reinforcement Learning Doesn't Work Yet — Systematic critique of deep RL limitations and failure modes
- DeepSeek - A Wake-Up Call For US Higher Education — DeepSeek's rise reflects China's STEM education investment advantage
- Early Evidence of Vibe-Proving with Consumer LLMs: A Case Study on Spectral Region Characterization with ChatGPT-5.2 (Thinking) — LLM-assisted iterative mathematical proof via generate-referee-repair pipeline
- Employment for computer programmers in the U.S. has plummeted to its lowest level since 1980—years before the internet existed | Fortune — AI correlates with historic drop in US programming jobs
- End of an Era at Google DeepMind Hints at New Future for AI — Last "Attention Is All You Need" author departs Google, ending transformer paper era
- Even experts are surprised by AI's latest 'vibe-mathing' advance — Amateur uses GPT-5.4 Pro to solve 60-year-old Erdős primitive sets problem
- Exhausted man defeats AI model in world coding championship — Human programmer narrowly beats OpenAI model at world coding championship
- Fable #6: The Return of the King — Anthropic Fable 5 restored after government export-control blip
- Fine-tuning ChatGPT: Surpassing GPT-4 Summarization Performance — 63% Cost Reduction and 11x Speed Enhancement using Synthetic Data and LangSmith — Fine-tuned ChatGPT beats GPT-4 summarization at 63% lower cost and 11x faster using chain-of-density synthetic data
- For Some Autistic People, ChatGPT Is a Lifeline — Autistic people using ChatGPT for social scripting and communication support
- Forget ChatGPT Canvas — I just tried Gemini Canvas and I'm floored by the difference — Gemini Canvas outperforms ChatGPT Canvas for AI-assisted writing feedback
- From Google To Nvidia, Tech Giants Have Hired Hackers To Break AI Models — AI red teams at major tech companies probing model vulnerabilities
- GPT-5.6: The System Card — Review of OpenAI GPT-5.6 Sol/Terra/Luna system card
- Generative AI needs tools to avoid copyright infringement, Databricks' Naveen Rao says — or more companies could meet Napster's fate — Generative AI copyright risk parallels Napster; open-source training on proprietary data as solution
- Generative AI: A Creative New World — Sequoia's 2022 thesis on generative AI market opportunity and waves
- George Carlin Estate Sues Creators of AI-Generated Comedy Special in Key Lawsuit Over Stars' Likenesses — First estate lawsuit over AI-generated deceased celebrity likeness and voice
- GitHub - daviddao/awful-ai: Awful AI is a curated list to track current scary usages of AI - hoping to raise awareness — Curated list cataloguing harmful and unethical real-world AI deployments
- GitHub - tatsu-lab/stanford_alpaca: Code and documentation to train Stanford's Alpaca models, and generate the data. — Instruction-following LLaMA model fine-tuned via self-instruct data
- Godfather of Artificial Intelligence" Geoffrey Hinton on the promise, risks of advanced AI — Geoffrey Hinton 60 Minutes interview warning of AI existential and societal risks
- Google DeepMind boss hits back at Meta AI chief over 'fearmongering' claim — Hassabis vs LeCun debate on AI safety, regulation, and open-source control
- Google's hidden AI diversity prompts lead to outcry over historically inaccurate images — Google Gemini paused after hidden diversity prompt injection produced historically inaccurate images
- Harvard '21 grad says Gen Z just uses A.I. to do their homework — Gen Z uses ChatGPT for homework completion rather than genuine learning
- Heuristics for lab robotics, and where its future may go — Three ideological camps driving wet-lab automation future
- How A.I. Is Changing the Way the World Builds Computers (Published 2025) — AI is fundamentally rebuilding computing hardware, data centers, and energy systems
- How ChatGPT is changing the way cybersecurity practitioners look at the potential of AI — ChatGPT dual-use cybersecurity capabilities surprise skeptical security researchers
- How Claude Helps Me Manage My Calendar (but ChatGPT Stumbles!) — Claude processes ICS calendar files reliably where ChatGPT fails context limits
- How I Used DALL·E 2 to Generate The Logo for OctoSQL — Iterative DALL-E 2 prompting and editing workflow for logo generation
- How Much of the World Is It Possible to Model? — Limits and nature of mathematical models from climate to LLMs
- How Peter Thiel's Relationship With Eliezer Yudkowsky Launched the AI Revolution — Thiel-Yudkowsky-Altman network origins of DeepMind and OpenAI
- How Smart is ChatGPT? — GPT-4 vs GPT-3.5 exam-percentile benchmark comparison
- How To Delete Your Data From ChatGPT — Guide to ChatGPT data deletion rights and privacy controls amid regulatory scrutiny
- How the Foundation Model Transparency Index Distorts Transparency — EleutherAI critique of Stanford FMTI biasing transparency toward corporate services
- How to Build An AI Agent with Function Calling and GPT-5 | Towards Data Science — Building a web-search AI agent via GPT-5 function calling
- How to design logos using DallE 3 - Complete guide — Guide to using DALL-E 3 for logo generation with prompt considerations
- Hugging Face clones OpenAI's Deep Research in 24 hours — Open-source research agent replicates OpenAI Deep Research in one day
- I'm running a 120B local LLM on 24GB of VRAM, and now it powers my smart home — Running gpt-oss-120b locally on 24GB VRAM via MoE
- IDEcline: How the world's most powerful coding tools became second-class citizens overnight — Three-wave shift from IDE-centric to agent-control-plane software development
- Inside the lucrative, disturbing world of humans training AI chatbots — Data annotators' pay, conditions, and experiences training AI models at Scale AI and others
- Intuitive RL: Intro to Advantage-Actor-Critic (A2C) | HackerNoon — Intuitive narrative introduction to the Advantage-Actor-Critic RL algorithm
- Is AI a danger to humanity or our salvation? — Hinton, LeCun, and Bengio split on AI existential risk after ChatGPT
- Jim Keller's startup rebrands Atomic Semi as Fab2, moves to Texas — Startup mass-producing small "fab fab" semiconductor factories
- Leaked org chart reveals the 58 top leaders and engineers at Google DeepMind — Google DeepMind org structure after 2023 Brain-DeepMind merger
- Llamar.ai: A deep dive into the (in)feasibility of RAG with LLMs — RAG product prototype built and shut down due to GPT-4 API cost infeasibility
- Look behind the curtain: Don't be dazzled by claims of 'artificial intelligence' | Op-Ed — Emily Bender op-ed demystifying AI as probabilistic pattern matching
- MCP doesn't move data. It moves trust — MCP as AI governance and intent-control layer over APIs
- MCP: The Missing Link Between AI Agents and APIs — MCP standardizes how AI agents access external APIs
- Meet AnythingLLM: A Full-Stack Application That Transforms Your Content into Rich Data for Enhanced Large Language Models LLMs Interactions — Open-source full-stack app for chatting with documents using LLMs
- Microsoft's relationship with OpenAI cracked when it hired Mustafa Suleyman, rival Marc Benioff says | TechCrunch — Microsoft-OpenAI partnership fracturing over competing AI ambitions
- Must-Have Prompt Engineering Skills for 2024 — In-demand prompt engineering skills, platforms, and tools
- New Data Shows Just How Badly OpenAI And Perplexity Are Screwing Over Publishers — AI search engines send 96% less referral traffic while massively increasing scraping
- New MIT Research Shows Spectacular Increase In White Collar Productivity From ChatGPT — MIT RCT finds ChatGPT makes white-collar workers 37% faster at writing tasks
- Noam Chomsky says A.I. is far from 'true intelligence' and ChatGPT is the 'banality of evil' | Fortune — Chomsky argues LLMs lack true reasoning and exhibit indifference to truth
- Once "too scary" to release, GPT-2 gets squeezed into an Excel spreadsheet — GPT-2 fully implemented in Excel spreadsheet for LLM education
- One of the three 'godfathers of A.I.' feels 'lost' because of the direction the technology has taken | Fortune — AI pioneers Bengio and Hinton warn of existential risk; LeCun dissents
- Open-Source AI Is Uniquely Dangerous — Unsecured open-source AI uniquely risky because safety features cannot be re-patched
- Open-source AI matches top proprietary model in solving tough medical cases — Open-source Llama 3.1 405B matches GPT-4 on clinical diagnostic reasoning
- OpenAI CEO Sam Altman says ChatGPT would have passed for an AGI 10 years ago — Sam Altman on shifting AGI definition, hallucinations, and AI training data consent
- OpenAI Offers A New Policy Blueprint — OpenAI's federal frontier-AI safety framework blueprint
- OpenAI Publishes GPT Prompt Engineering Guide — OpenAI's six-strategy guide for eliciting better GPT-4 responses
- OpenAI Says It's "Over" If It Can't Steal All Your Copyrighted Work — OpenAI lobbies White House to expand fair use for AI training data
- OpenAI executives say releasing ChatGPT for public use was a last resort after running into multiple hurdles — and they're shocked by its popularity — OpenAI's surprise at ChatGPT's viral adoption and CEO's AI risk warnings
- OpenAI is throwing everything into building a fully automated researcher — OpenAI's North Star automated AI researcher agent
- OpenAI says there's only a small chance ChatGPT will help create bioweapons — OpenAI self-study finds GPT-4 gives marginal bioweapon research uplift
- OpenAI to acquire Neptune — OpenAI acquisition of neptune.ai ML experiment tracking platform
- OpenAI to acquire Neptune — OpenAI acquires Neptune experiment tracking platform for training stack
- OpenAI wasn't expecting Sora's copyright drama — Sora launch copyright backlash and OpenAI policy reversal at DevDay 2025
- OpenAI's 'Pay As You Go' Is the Best Way to Use ChatGPT — OpenAI pay-as-you-go API plan cheaper than ChatGPT Plus subscription
- OpenAI's board has fired Sam Altman — Hacker News thread on OpenAI board firing Altman
- Oxford shuts down institute run by Elon Musk-backed philosopher — Oxford closes Future of Humanity Institute after 19 years amid scandals
- People Are Using A 'Grandma Exploit' To Break AI - Kotaku — Roleplay persona prompts bypass AI safety guardrails
- Perspectives on the Social Impacts of Reinforcement Learning with Human Feedback — Social and ethical impacts of RLHF across seven societal dimensions
- Practical Guide to Task Automation using ChatGPT & Python — Using ChatGPT and Python APIs to automate common knowledge-work tasks
- Re-implementing LangChain in 100 lines of code — Hacker News debate over LangChain's abstraction value
- Reddit to charge for API access; CEO blames A.I. | Fortune — Reddit monetizing API access to prevent free AI training data extraction
- Reinforcement Learning algorithms — an intuitive overview — Survey of model-free and model-based RL algorithm families
- Sam Altman enters his power era — Sam Altman reinstated as OpenAI CEO after failed board ouster
- Samsung workers made a major error by using ChatGPT — Samsung engineers leaked trade secrets via ChatGPT input data retention
- Showrunner wants to turn you into a happy little content prompter for the 'Netflix of AI' — Fable's Showrunner AI platform for prompt-generated shows
- Silicon Valley bets big on 'environments' to train AI agents | TechCrunch — RL environments emerge as critical training infrastructure for capable AI agents
- Space Force Gets Scared, Pauses All Use of Generative AI — US Space Force bans generative AI tools on government devices over data security concerns
- Teaching with AI — OpenAI guide for educators using ChatGPT in classrooms
- The 'Godfather of AI' Has a Hopeful Plan for Keeping Future AI Friendly — Geoffrey Hinton's views on LLM risks and analog computing as AI safety mitigation
- The AI Agent Era Requires a New Kind of Game Theory — CMU researcher on agent security risks and multi-agent game theory
- The Batch | DeepLearning.AI | AI News & Insights — DeepLearning.AI weekly AI news and insights newsletter homepage
- The Company Behind Stable Diffusion Appears to Be Crumbling Into Chaos — Stability AI leadership exodus and CEO credibility crisis in mid-2023
- The Death of the "Everything Prompt": Google's Move Toward Structured AI — Google Interactions API for stateful agentic AI
- The Doomsday Invention — Nick Bostrom's superintelligence thesis and AI existential risk debate
- The False Promise of Imitating Proprietary LLMs — Finetuning open models on ChatGPT outputs mimics style but not factuality or capability
- The Generative AI Copyright Fight Is Just Getting Started — Legal debate over AI training on copyrighted works and fair use
- The Goopification of AI — AI chatbots colonizing the self-help genre via probabilistic text assembly
- The Pentagon says AI is speeding up its 'kill chain' | TechCrunch — AI accelerating Pentagon kill chain planning amid usage policy tensions
- The Singular Value Decompositions of Transformer Weight Matrices — SVD of GPT-2 weight matrices reveals interpretable semantic directions
- The author of SB 1047 introduces a new AI bill in California | TechCrunch — California SB 53 creates AI whistleblower protections and public compute cluster
- The challenges of reinforcement learning from human feedback (RLHF) - TechTalks — RLHF limitations across feedback, reward modeling, and policy
- The founder behind ChatGPT once gave a lecture on how to launch a startup. It's still required viewing 9 years later. — Sam Altman's 2014 Stanford startup lecture key principles still relevant
- The pathway to Transformers — Technical walkthrough of architectural evolution from RNNs to Transformers
- The second wave of AI coding is here — Next-gen AI coding agents targeting functional correctness via process data
- These ex-Apple employees are bringing AI to the desktop — Ex-Apple founders launch startup for LLM-powered desktop OS reimagining
- Three Labs With a Plan and A Memorandum — Trump AI memorandum plus OpenAI's benefit-everyone plan
- Three weeks after acquiring Windsurf, Cognition offers staff the exit door | TechCrunch — Cognition acquires Windsurf then immediately offers buyouts to staff
- Training Your Own LLM using privateGPT — Running a private local LLM on sensitive data without cloud exposure
- U.S. regulators warn they already have the power to go after A.I. bias — and they're ready to use it — Four US agencies assert existing legal authority to enforce against AI bias
- USAF Official Says He 'Misspoke' About AI Drone Killing Human Operator in Simulated Test — Viral USAF AI drone kills-operator story was a hypothetical thought experiment
- Unauthorized "David Attenborough" AI clone narrates developer's life, goes viral — GPT-4V plus ElevenLabs creates unauthorized celebrity voice clone
- Unbabel says its new AI model has dethroned OpenAI's GPT-4 as the tech industry's best language translator — TowerLLM 7B/13B model edges GPT-4o on multilingual translation benchmarks
- Using ChatGPT as a technical writing assistant — ChatGPT as iterative technical writing draft tool
- WSJ Article Claiming China Has Matched Anthropic Is Obvious Nonsense — Debunking WSJ headline that China matched Mythos on cyber
- What Really Made Geoffrey Hinton Into an AI Doomer — Hinton's reasons for leaving Google to warn about accelerating AI risk
- What was 60 Minutes thinking, in that interview with Geoff Hinton? — Gary Marcus rebuttal annotating 60 Minutes Hinton AI interview
- White House Will Ad Hoc Decide Who Can Individually Access GPT-5.6 — Critique of ad hoc White House frontier-model access regime
- Who Should Stop Unethical A.I.? — Debate over ethics review mechanisms for AI research publication
- Why some college professors are adopting ChatGPT AI as quickly as students — ChatGPT disruption of higher education from professor perspective
- Will Transformers Take Over Artificial Intelligence? | Quanta Magazine — Transformers expanding from NLP to vision, generative, and multimodal AI tasks
- Wix will let you build an entire website using only AI prompts — Wix AI Site Generator builds full websites from text prompts using ChatGPT and DALL-E
- Worldcoin, co-founded by Sam Altman, is betting the next big thing in AI is proving you are human | TechCrunch — Iris-scanning proof-of-personhood system as defense against AI-indistinguishable bots