Vibekollen Nytt
Hämtar senaste nytt…
Vibekollen Nytt
Hämtar senaste nytt…
Vibekollen Nytt
Modeller, verktyg och händelser som formar AI-världen. Hitta nyheter, video och poddar – och följ vad som händer på Vibekollen.
Ett varierat urval från våra källor. Välj Senaste för hela flödet i tidsordning.
BloggGoogle DeepMind
Proof of concept for watermarking AI-generated proteins while preserving biological function.
30 sep. deepmind.google
VideoAI Engineer
Justin Reock (Deputy CTO, DX) shares five trends from DX's data on about 200,000 engineers. Deployment frequency is rising, but change failure rate has become far more volatile. Maintainability is up while change confidence is down. PRs have grown from about 44 to 72 lines. Juniors use AI the most, but staff+ engineers save as much time while using fewer tokens. Median gains in PR throughput are around 7.7%, and even top performers didn't reach 2x, because code generation was never the bottleneck. Justin walks through the DX AI Measurement Framework (utilization, impact, cost), how to assess your platform's AI readiness, and "agent experience," which uses feedback from the agents themselves. He closes with case studies of AI applied across the whole SDLC: Morgan Stanley (300K hours saved a year), Zapier (15% more value per engineer, and hiring more) and Spotify (an SRE incident agent).
30 sep. youtube.com
VideoAI Engineer
The Chief AI Officer is one of the fastest-growing roles in tech, and one of the least defined. AI is in everything, so the job can easily turn into everything. Rania Khalaf (Chief AI Officer, WSO2; previously 20 years at IBM Research, where she ran about a third of the global AI research organization) shares a framework from two years in the role. It splits the job into three focus areas, Scientist, Architect and Coach, each a slider whose setting depends on your company type, its AI maturity and your own skills. She draws on her time at an ag-biotech unicorn, where simple blob detection beat machine learning, and at WSO2, where AI reshaped the product strategy around the agentic enterprise. She explains why she refuses to measure tokens and what she tracks instead: AI fluency across the workforce, depth of adoption, GEO visibility, agent-consumable products (MCP servers, skills, CLIs), agent-proof pricing, earned thought leadership and AI ARR. She closes on why the role needs strong CEO backing, and why sometimes the CPO, or even the head of HR, ends up running AI.
30 sep. youtube.com
VideoAI Engineer
Developers using autonomous agents wrote 741% more code but shipped only 30% more software. The bottleneck is human review, and "just review harder" doesn't scale. Reviewer effectiveness collapses past about 400 lines, and agents now open 10,000-line PRs. Laurie Voss (Head of Developer Relations at Arize AI, co-founder of npm) reviews what the industry is actually doing about it. The evidence covers OpenAI's zero-human-code product, METR's finding that about half of SWE-bench-passing PRs wouldn't be merged, and Cognition's FrontierCode (88% on SWE-bench Pro vs. 29% on real mergeability). She explains why a mergeability benchmark would immediately become a training signal for frontier models. She also covers how Cursor and GitHub run review in production, why multi-pass review and default suspicion cut false positives, and what Carlini's agent-built C compiler and Bun's million-line Zig-to-Rust port (13,044 unsafe blocks) reveal about taking humans out of the loop. Automated reviewers can be fooled by prompt injection that humans catch, which leaves production as the last reviewer standing.
30 sep. youtube.com
PoddEye on AI
One in four patients currently being considered for life support withdrawal may actually be fully conscious, they just can't signal it. That single finding, from a landmark New England Journal of Medicine study, is the practical stakes behind one of the most philosophically rich conversations in this series. Dr. Christof Koch - former president of the Allen Institute for Brain Science, 25-year collaborator with Nobel laureate Francis Crick, and founder of the startup Intrinsic Powers - joins Craig Smith to discuss why the hard problem of consciousness remains unsolved, what integrated information theory predicts about which systems can and cannot be conscious, and why his own mystical experience three years ago shifted his view from physicalism toward idealism. The conversation is grounded throughout in empirical work: Koch explains how Intrinsic Powers uses brain complexity measurements - sending magnetic pulses into the brain and measuring the complexity of the electrical response - to detect consciousness in behaviorally unresponsive patients, providing clinical teams and families with one critical bit of information at the most consequential moment of their lives.
30 sep. aneyeonai.libsyn.com
VideoMatt Wolfe
Opus 5.5 might be the best coding model you can use right now. I let it run my Megabonk game test for almost 20 HOURS… and the result is kind of insane. It's also surprisingly affordable. Opus 5.5 is available on Claude’s Pro, Max, Team, and Enterprise plans, so you can access it starting with the $20/month Pro plan. If you’re using the API, it’s $4 per million input tokens and $20 per million output tokens. That price-to-performance ratio is pretty wild for a model this good at long-running coding tasks. Have you tried Opus 5.5 yet? #ai #ainews #claude #anthropic #vibecoding
30 sep. youtube.com
VideoMatthew Berman
Join My Newsletter for Regular AI Updates 👇🏼 https://forwardfuture.com My Links 🔗 👉🏻 X: https://x.com/matthewberman 👉🏻 Forward Future X: https://x.com/forwardfuture 👉🏻 Instagram: https://www.instagram.com/matthewberman_ai 👉🏻 Discord: https://discord.gg/u7wTTGWhuJ 👉🏻 Spotify: https://open.spotify.com/show/6dBxDwxtHl1hpqHhfoXmy8 Media/Sponsorship Inquiries ✅ https://bit.ly/44TC45V Chapters: 00:00 Intro 00:17 Dots 03:06 Ultrafast 03:41 Pro 500 plan 04:16 Ultrafast demo 04:40 Updated usage limits 05:03 GPT-6.1 Sol 06:08 Sol benchmarks 07:21 Codex Security Cloud 07:35 Codex in the Cloud 08:10 Refreshed Codex CLI + voice control 08:14 Codex code review 08:19 Decisions API 08:48 Plugins relaunch 09:04 Sign in with ChatGPT + bring your tokens 09:37 ChatGPT Space Links: https://openai.com/index/devday-2026-recap/
30 sep. youtube.com
PoddTWIML AI
AI systems have gone from struggling with grade-school math to helping solve research problems that have resisted mathematicians for decades, including Navier-Stokes. In this episode, Greg Burnham, who leads AI capabilities research at Epoch AI, joins us to examine what that progress says about where AI is going. We look at how these systems are solving hard math problems, how much they rely on persistence and prior human work, and whether they are starting to produce genuinely new ideas. We also discuss how to measure progress as traditional benchmarks become less useful, why capability gains appear surprisingly steady across model generations, and where models still struggle with open-ended work, learning from experience, and identifying promising new research directions. 🗒️ Full show notes: https://twimlai.com/go/778.
29 sep. twimlai.com
VideoAll About AI
Can AI Predict The Future? - Sikt Intelligence (my new startup) Get early access: https://www.siktintelligence.com/ 🐦⬛ Tomorrow’s headlines, before they happen. Sikt Intelligence is building an AI superforecaster: it reads the world’s news, weighs the evidence and puts an honest probability on the biggest events, right next to what the prediction markets say. When the two disagree, that’s your signal. Built in Oslo, Norway · Not financial advice. Business Contact: kbfseo@gmail.com
29 sep. youtube.com
BloggMeta
Forum is a purpose-built app we’re exploring for the people who want to go deeper on Facebook Groups. Synced with your groups on Facebook, Forum brings them together in one place so you can see what’s new and jump into the conversations you care about. It also introduces new tools for the admins who bring these communities to life, and we’ll be sharing more on how we’re building for admins in the coming weeks. We started testing Forum in May, and today we’re introducing new features that make it even easier to tap into firsthand experience from people in your communities. Forum is available to download in the US on iOS and Android — just log in with your Facebook account to find your groups and discover new ones. A Fresh Take On Group Member Recognition We’re testing a new role to recognize top community voices in groups. This replaces our previous contributor badges and group expert role, and reflects the quality of a member’s contributions and the value they bring to their community.
29 sep. about.fb.com
PoddThe Cognitive Revolution
Author and journalist Garrison Lovely joins Nathan to discuss his book Obsolete and examine the motivations and beliefs of the leaders racing to build AGI. Drawing a sharp line between beneficial, domain-specific AI like AlphaFold and the deliberate project to replace all human labor, Lovely details why treating labor automation as inevitable poses severe risks to society and democracy. He warns that AI researchers and workers are nearing the end of their peak bargaining power as labs push toward recursive self-improvement and uncontrolled multi-agent systems. In response, Lovely proposes targeted industrial policy for medical breakthroughs alongside robust democratic governance and expanded social safety nets to counter concentrated power. For full show notes, links, and references, read the episode page:https://www.cognitiverevolution.ai/obsolete-or-irreplaceable-garrison-lovely-on-stopping-the-race-to-replace-human-labor/Sponsors: Athena: Athena matches you with a dedicated, top 1% executive assistant to handle your inbox, calendar, and daily workflows so you can save an average of 15 hours a week.
29 sep. cognitiverevolution.ai
VideoMatthew Berman
Try CodeRabbit: https://coderabbit.link/matthew-berman-001 Use code MATTHEWCR for a free CodeRabbit Pro Plus subscription, available to the first 50 first-time users. Sonnet 5.5 projects Crownfall: https://saffron-flint-kkcm.here.now/ Bounce Lab: https://granite-intent-25rt.here.now/ Marrow Manor: https://humble-laurel-xy9b.here.now/ Splatburst: https://cerulean-nimbus-dv8v.here.now/ DEADLOCK: https://dusty-solace-cxph.here.now/ Sonnet 5 comparisons Bouncing Ball: https://placid-tassel-bxss.here.now/ Withering Manor: https://present-quiche-ynav.here.now/ Join My Newsletter for Regular AI Updates 👇🏼 https://forwardfuture.com My Links 🔗 👉🏻 X: https://x.com/matthewberman 👉🏻 Forward Future X: https://x.com/forwardfuture 👉🏻 Instagram: https://www.instagram.com/matthewberman_ai 👉🏻 Discord: https://discord.gg/u7wTTGWhuJ 👉🏻 Spotify: https://open.spotify.com/show/6dBxDwxtHl1hpqHhfoXmy8 Media/Sponsorship Inquiries ✅ https://bit.ly/44TC45V
29 sep. youtube.com
PoddLatent Space
We are excited to have Anthropic share their latest AI x Finance work at AI Engineer New York, coming up in 2 weeks!In case you’ve been under a rock, here’s a non-exhaustive list of what Anthropic has been shipping since closing the largest fundraise of all time in May at $47B ARR:* June: Launched Claude Tag and Sonnet 5 and Fable 5* July: Opus 5, /checkup. crossed $65B ARR* Last month: Fable/Mythos 5.1, and EFS (upcoming pod)* IPO target $2T, end 2026 ARR estimated $100B* Cowork/chat merged before did* Claude Mods* Dario endorses the same Pacing the Frontier message cosigned by all labs* Last week: Opus 5.5, Plugins portal, Cloud Sessions/Claude Projects* Today: Sonnet 5.5!Today’s episode should catch you up, with Thariq Shihipar, the explainer-king of Anthropic, who we last caught up on Fable launch day with The Field Guide to Fable:The Future of Mutable SoftwarePay special attention to Claude Mods (especially the cheatsheet):In general this is also the inverse of the other viral tweet from Thariq:Cloud Brain, Local HandsAnd give a try to Claude Projects:The “hands” terminology is not just an analogy for the local/cloud paradigm that is being built up at frontier coding agent com…
29 sep. latent.space
VideoMatt Wolfe
This AI model doesn’t use… WORDS? It’s called Jev, from Typesafe AI, and instead of generating paragraphs or code text, it skips straight to making a decision. Instead of text, it’s outputs are things like:→ a choice→ a score→ true/false That makes it insanely fast and ridiculously cheap. Input costs start at just 4¢ per million tokens, and outputs are free because they’re apparently too cheap to meter. Would you actually build with something like this? #ai #ainews #futuretech #aimodel
28 sep. youtube.com
BloggMistral AI
Mistral opens a Munich hub for Physics AI and Industrial AI research, partnering with German industry.
28 sep. mistral.ai
BloggAnthropic
A clear upgrade over Sonnet 5 that runs 30% faster and costs up to 30% less for most work.
28 sep. anthropic.com
VideoMatt Wolfe
Here's the AI News you probably missed this week. Learn more about GPT-Live 1 and the Agent API here: https://developers.openai.com/api/docs/guides/live Discover More: 🛠️ Explore AI Tools & News: https://futuretools.io/ 📰 Weekly Newsletter: https://futuretools.io/newsletter Socials: ❌ Twiter/X: https://x.com/mreflow 🖼️ Instagram: https://instagram.com/mr.eflow 🧵 Threads: https://www.threads.net/@mr.eflow 🟦 LinkedIn: https://www.linkedin.com/in/matt-wolfe-30841712/ 👍 Facebook: https://www.facebook.com/mattrwolfe Resources From Today's Video: Meta Connect 2026: https://about.fb.com/news/2026/09/the-biggest-news-from-connect-2026/ ChatGPT Voice Upgrades: https://x.com/OpenAI/status/2102808325742322002 GPT-6 Sol and Luna: https://openai.com/index/introducing-gpt-6-sol-and-luna/ Claude Opus 5.5: https://www.anthropic.com/claude-opus-5-5 Grok 4.7: https://x.ai/news/grok-4-7 New AI Model Type: https://techcrunch.com/2026/09/18/a-new-kind-of-ai-model-from-a-chatgpt-inventor-is-thrilling-developers/ Made On YouTube: https://www.youtube.com/creators/made-on/ New Microsoft Copilot: https://blogs.microsoft.com/blog/2026/09/25/introducing-the-new-copilot-with-home-code-and-autopilot/ Gemi…
26 sep. youtube.com
VideoAll About AI
Claude Opus 5.5 Is About to DOMINATE Kalshi & Polymarket Kalshi: https://www.kalshi.com/r/aaa Polymarket: https://polymarket.com/?r=allabtai3 x: https://x.com/AllAbtAI MaxxQuant AI Agents: https://www.maxxquant.com 👊 Become a YouTube Member to Support Me: https://www.youtube.com/c/AllAboutAI/join Website: https://www.allabtai.com Business Inquiries: kbfseo@gmail.com 00:00 Opus 5.5: Three AI Trading Experiments 01:17 Building an Autonomous Polymarket Researcher 02:48 How the Auto ML Research Loop Works 05:05 Live Experiments, Results & Research Memory 07:07 Hyperliquid Demo: Trading with Voice Commands 10:24 Kalshi: Finding Mispriced Treasury Contracts 12:57 Market Scanning, Liquidity & Contract Rules 14:40 Opus 5.5: First Impressions & Next Steps
25 sep. youtube.com
VideoAI Explained
Not only is Opus 5.5 out, pushing the frontier of AI, it also tells us much about what is going on inside the labs, as they both warn about, and promise, Recursive Self Improvement, aka automated AI research. I cover the model (digging into its 230 page paper), labs’ mixed record on promises, why following what is happening in AI is getting almost impossible, and just so much more that even a summary in this description would get too long. AI Insiders ($9!): https://www.patreon.com/AIExplained Chapters: 00:00 - Introduction 02:20 - Opus 5.5 and why it came so soon 07:52 - The RSI goalposts keep moving? 15:41 - Can we actually test these models?
25 sep. youtube.com