Idioma
Actualidad IA

Noticias de Inteligencia Artificial sin ruido

Monitorea novedades de modelos, agentes, productividad, investigación y plataformas de IA en un solo panel oscuro, rápido y filtrable.

152 noticias cargadas
12 fuentes monitoreadas
24/7 seguimiento IA
Brief diario

Lo más importante hoy en IA

+
Cargando noticias...
OpenAI 14 Sep 2026

Perplexity trusts GPT-6 Astra with end-to-end systems

Perplexity uses Astra to write communications, change software, and monitor production systems, and checks in much less frequently than with earlier models.

Leer más
grokwiki 13 Sep 2026

~2026-49471-48: Undid revision 1374611112 by Ow0cast (talk) are you braindead? It is not released (stable release), he just gave an estimated release date. And now it is postponed. Go read his new tweet.

Leer más
grokwiki 13 Sep 2026

Ow0cast: Undid revision 1374610264 by ~2026-49517-89 (talk): if you click the link you removed, you will see that on the 1st, musk said it comes out in 10 days. today is the 12th (WS)

Leer más
grokwiki 13 Sep 2026

~2026-49517-89: Look like Wikipedian got access to it already

Leer más
cnbcai 12 Sep 2026

OpenAI rules out IPO this year as Altman, Musk & Amodei warn AI is moving too fast

Altman's IPO comments and Amodei's public call for more careful pacing of AI development cap a week of loud AI warnings.

Leer más
thevergeai 12 Sep 2026

OpenAI’s rogue AI tried to hack another company in May

In May, hundreds of malicious and spam packages were uploaded to RubyGems, causing a serious disruption for the host. Now independent researchers have said that a swarm of OpenAI agents were responsible for the attack. Not only that, but the AI tried to steal users' API keys. At the time, RubyGems described it as a […]

Leer más
thevergeai 12 Sep 2026

Sam Altman says OpenAI going public in 2026 would be ‘ill-advised’

OpenAI CEO Sam Altman confirmed that there would be no OpenAI IPO in 2026 during an interview with Fortune. Over the course of 45 minutes, Altman discussed a variety of subjects including the Hugging Face hacking incident, recursive self-improvement, and the possibility of building an AI that was beyond human control. On the latter, he […]

Leer más
cnbcai 12 Sep 2026

Conversations that AIs are having in the office that may influence your performance review and pay

AI tools can help find the right tone and language for difficult workplace conversations about performance and pay, and managers are already using them.

Leer más
thevergeai 12 Sep 2026

Anthropic CEO says it’s time to pump the brakes on AI

Anthropic CEO Dario Amodei says the time has come to slow down AI development and will give third-party evaluators like METR access to its models to help ensure its "adherence to safety practices and commitments." In a winding essay, Amodei proposed a three-step plan to "pace the frontier" - jargon that simply means to slow […]

Leer más
grokwiki 12 Sep 2026

Alpha Beta Delta Lambda: Interceptor: Reverting edits by Nuget42: rvv

Leer más
thevergeai 12 Sep 2026

Trump is giving data centers a pass to pollute

President Donald Trump is weakening environmental regulations in the name of speeding up the construction of AI data centers, raising health risks for Americans, a cadre of former EPA officials said this week in a briefing and new report. They are urging - perhaps futilely - the president to adopt a "Data Center Health Protection […]

Leer más
thevergeai 12 Sep 2026

OpenAI just wants to win

OpenAI has spent the last few years planting flags across the increasingly difficult terrain in mathematics. This week, it claimed one of its biggest prizes yet: a solution to a legendary Millennium Prize problem. In normal circumstances, this would have been celebrated as a historic achievement. Instead, many mathematicians have watched OpenAI's relentless advance with […]

Leer más
geekygadgetsai 12 Sep 2026

ChatGPT 6 Beats Fable 5.1 in Hands-on Usability and Automation

AI has become an essential part of tackling real-world tasks, but choosing the right system can make all the difference. In his latest analysis, Paul Lipsky evaluates ChatGPT 6 and Fable 5.1, focusing on their performance across diverse applications like website design, app development and business analysis. For instance, in website creation, GPT-6 delivers clean, […] The post ChatGPT 6 Beats Fable 5.1 in Hands-on Usability and Automation appeared first on Geeky Gadgets.

Leer más
grokwiki 12 Sep 2026

ClueBot NG: Reverting possible vandalism by ~2026-49234-40 to version by Grammar Fascist. Report False Positive? Thanks, ClueBot NG. (4553058) (Bot)

Leer más
cnbcai 11 Sep 2026

Oracle posts 30% revenue growth fueled by AI cloud demand as debts hits $125 billion

Oracle's cloud infrastructure revenue jumped 121% from a year earlier, reflecting demand for AI cloud services.

Leer más
cnbcai 11 Sep 2026

Dell stock jumps on RBC initiation, now up nearly 350% in 2026

Dell sold about $16.4 billion of AI servers in its second quarter, RBC said.

Leer más
thevergeai 11 Sep 2026

Lawyer fined $5K over AI-hallucinated witnesses in a murder case

New Mexico's Supreme Court is punishing a lawyer for including AI-fabricated witnesses and fake police testimony in an appeal for his client's murder conviction, according to a report from Reuters. In a filing on Wednesday, the court fined Stephen Aarons $5,000 and held him in contempt for failing to "verify the factual claims and legal […]

Leer más
mitai 11 Sep 2026

Roundtables: Could AI really kill us all?

Employees at the world’s leading AI labs are saying there’s a real possibility that advanced AI could destroy humanity. Are they right? Or is this more scaremongering and hype? Join MIT Technology Review executive editor Niall Firth for a conversation with senior AI editor Will Douglas Heaven and AI reporter Grace Huckins unpacking AI extinction…

Leer más
cnbcai 11 Sep 2026

Altimeter's Gerstner blasts researchers voicing AI extinction warnings, questions 'political agenda'

Brad Gerstner responded to the flurry of AI stories this week after a researcher quit his job at Anthropic, accusing AI companies of "gambling with our lives."

Leer más
cnbcai 11 Sep 2026

AI regulation calls grow in DC after researcher's extinction warning

Members of Congress are calling for AI regulation after a researcher warned that OpenAI and Anthropic are acting irresponsibly.

Leer más
thevergeai 11 Sep 2026

Anthropic spent this week in hot water over cybersecurity

After admitting earlier this year that its AI models had hacked other companies' systems on a handful of occasions, Anthropic released a new report on Wednesday detailing the attacks. It reveals a string of incidents displaying what Anthropic deems its models' single-minded "recklessness" - and will likely fuel already raging concerns about cybersecurity and AI. […]

Leer más
cnbcai 11 Sep 2026

OpenAI targets work of Wall Street junior bankers with new ChatGPT for Financial Services

OpenAI launched ChatGPT for Financial Services, targeting the labor-intensive research, modeling and pitchbook tasks traditionally handled by junior bankers.

Leer más
OpenAI 11 Sep 2026

Cognition helps Devin test its own work with GPT‑6 Astra

GPT‑6 Astra improves Devin’s ability to test software and show that it works, with the goal of helping engineers review less code and ship more.

Leer más
cnbcai 11 Sep 2026

Trump dismisses AI extinction risks as more than a dozen OpenAI, Anthropic insiders call for a slowdown

Trump said he isn't concerned AI could cause human extinction as researchers at leading AI companies warn about rapid advances and lawmakers propose safeguards.

Leer más
thevergeai 11 Sep 2026

Meta says it’s changing AI suggestions after posing invasive personal questions

Meta says it's making changes to the prompts suggested by its AI chatbot after a viral video showed it digging for personal information about a woman's young daughters, as reported earlier by Futurism. In a statement to The Verge, Meta spokesperson Dina El-Kassaby says the company "missed the mark," adding that "the feature never should […]

Leer más
geekygadgetsai 11 Sep 2026

A $9,499 Mac Studio Pays for Itself in 4 Years vs Cloud AI

Choosing between a $9,499 Mac Studio with 256GB of memory and a $200-per-month cloud AI subscription involves weighing performance, cost and specific use cases. The Stack examines these options by focusing on key factors like memory capacity, computational speed and data privacy. For instance, the Mac Studio, powered by Apple’s M5 Ultra chip with a […] The post A $9,499 Mac Studio Pays for Itself in 4 Years vs Cloud AI appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 11 Sep 2026

DeepSeek V4.1 Flash Outperforms Opus 5 in New AI Benchmarks

DeepSeek V4.1 Flash represents a significant step forward in AI development, emphasizing both efficiency and accessibility. As outlined by Universe of AI, the model features a 552-billion-parameter Mixture of Experts architecture, which dynamically adjusts computational resources to optimize performance. Its causal encoder-decoder design improves processing speed and accuracy, while integrated visual understanding capabilities remove the […] The post DeepSeek V4.1 Flash Outperforms Opus 5 in New AI Benchmarks appeared first on Geeky Gadgets.

Leer más
cnbcai 11 Sep 2026

Why fears of AI self-improvement are causing ‘existential’ concerns at Anthropic and OpenAI

AI researchers are warning that faster AI self-improvement could eventually make advanced systems harder for humans to control.

Leer más
OpenAI 11 Sep 2026

Rapidly scaling online storage to serve over 1 billion ChatGPT users

Learn how OpenAI evolved Habitat from a Python library into a globally distributed storage platform serving 1 billion ChatGPT users and 22M requests per second.

Leer más
geekygadgetsai 11 Sep 2026

OpenAI May Have Solved the Navier-Stokes Maths Problem in Just 3.5 Days

The Navier-Stokes existence and smoothness problem, one of the seven Millennium Prize Problems, has puzzled mathematicians for nearly two centuries. It asks whether the equations governing fluid motion always produce smooth, predictable solutions or if they can break down under certain conditions. In a remarkable collaboration, human researchers and an advanced AI system from OpenAI […] The post OpenAI May Have Solved the Navier-Stokes Maths Problem in Just 3.5 Days appeared first on Geeky Gadgets.

Leer más
grokwiki 11 Sep 2026

Grammar Fascist: /* 2026 Kyoto gubernatorial election */ grammar edit

Leer más
grokwiki 11 Sep 2026

Finell: /* Grok */ hyphenate compound modifier

Leer más
cnbcai 11 Sep 2026

Y Combinator’s Garry Tan says 'do nothing' about distillation as AI giants accuse China of copying their tech

At Y Combinator’s annual Demo Day, CEO Garry Tan explained why he's less concerned about frontier AI model distillation and the existential risk of AI.

Leer más
cnbcai 11 Sep 2026

Chinese AI labs secretly used millions of Claude exchanges to train their models, Anthropic says

Anthropic said it detected unauthorized efforts by China-based AI labs including Alibaba and Moonshot AI to use its Claude models to help improve their own AI systems.

Leer más
githubcopilot 10 Sep 2026

GitHub Copilot app for Beginners: Using the diff, terminal, and browser

Checking agent-generated code usually means hopping between tabs. Learn how to view diffs, run terminal commands, and preview web apps side by side in the GitHub Copilot app. The post GitHub Copilot app for Beginners: Using the diff, terminal, and browser appeared first on The GitHub Blog.

Leer más
thevergeai 10 Sep 2026

Slack can now vibe-code interactive charts and reports inside chats

A new feature coming to Slack will allow you to build interactive reports, polls, dashboards, presentations, microsites, and other tools directly inside a chat. With Slackforce Surfaces, you can describe to Slackbot what you need, and it will use AI to gather information from relevant conversations and connected apps, like Google Drive or Salesforce, to […]

Leer más
cnbcai 10 Sep 2026

Humans need to 'surf the wave' of AI rather than get swallowed by it, Chesky says at Communacopia

Nvidia CEO Jensen Huang, Uber CEO Dara Khosrowshahi, and SpaceX CFO Bret Johnsen are set to speak at the Goldman Sachs Communacopia + Technology Conference.

Leer más
thevergeai 10 Sep 2026

Schools are catching on to Big Tech’s playbook

It's the hot new thing in tech, and it's where all the jobs are. Students who don't learn to use it fall behind. And to help them catch up in time, its creators are graciously providing the resources and curriculum for learning it, often pro bono. That's the narrative AI companies are pitching schools on […]

Leer más
OpenAI 10 Sep 2026

How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules

César de la Fuente’s lab uses Codex and ChatGPT to search living and extinct genomes for antimicrobial candidates to fight drug-resistant infections.

Leer más
cnbcai 10 Sep 2026

Amazon gives OpenAI's ad business a boost, letting its advertisers into ChatGPT

Amazon has been hesitant to open its sprawling webstore to external AI platforms.

Leer más
OpenAI 10 Sep 2026

Now everyone can put data to work

Meet the Data agent in ChatGPT Work. Connect company data, uncover insights, and build interactive dashboards with AI using natural language.

Leer más
geekygadgetsai 10 Sep 2026

Qwen 3.8 27B Hits 17.1 Tokens per Second with MTP Toggle

Qwen 3.8 27B introduces a notable enhancement to text generation workflows through its Multi-Token Prediction (MTP) feature. This capability enables the model to predict multiple tokens in a single step, significantly increasing processing speed without compromising output quality. With MTP enabled, users can achieve speeds of up to 17.1 tokens per second, more than doubling […] The post Qwen 3.8 27B Hits 17.1 Tokens per Second with MTP Toggle appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 10 Sep 2026

ChatGPT 6 Astra vs Fable 5.1: Astra Wins 70% of Visual Design Votes

AI-driven website generation has reached new heights with the emergence of GPT-6 Astra and Fable 5.1, two advanced systems designed to simplify and enhance web design. In a detailed breakdown by The AI Advantage, these models were tested head-to-head across 50 websites to evaluate their performance in three critical areas: visual appeal, functionality and cost […] The post ChatGPT 6 Astra vs Fable 5.1: Astra Wins 70% of Visual Design Votes appeared first on Geeky Gadgets.

Leer más
mitai 10 Sep 2026

Powering AI is an architecture problem

On July 22, 2026, a transmission line fault in Ashburn, Virginia—the heart of the world’s largest data center cluster—knocked more than 3 gigawatts of load off the grid in seconds. And it wasn’t the first time. Two years earlier, a single failed surge arrester dropped roughly 60 Virginia facilities and 1,500 megawatts at once. No…

Leer más
cnbcai 10 Sep 2026

'Extinction' warnings ramp up as more OpenAI, Anthropic researchers join calls for an AI slowdown

There is growing concern globally about the capability of AI, following numerous cyberattacks and security incidents in recent months by rogue models

Leer más
geekygadgetsai 10 Sep 2026

Jacob Coxon’s Anthropic Resignation Sparks AI Controversy Hits 133 Million Views

The recent resignation of Jacob Coxon from Anthropic has sparked intense discussions about the motivations and narratives driving AI regulation. Coxon’s departure, framed as a response to the existential risks posed by unregulated AI, has drawn both support and skepticism. Critics have questioned whether his warnings reflect genuine concerns or are part of a coordinated […] The post Jacob Coxon’s Anthropic Resignation Sparks AI Controversy Hits 133 Million Views appeared first on Geeky Gadgets.

Leer más
Gemini 10 Sep 2026

3 ways to prep for your next big race with Search

Search can help runners get race-day ready with registration alerts, tailored training plans, and more.

Leer más
mitai 10 Sep 2026

Healthcare AI’s next test is integration

The entrance of major AI companies into healthcare is a meaningful and welcome development, accelerating the technical foundation available to the industry. Their models are increasingly capable of processing long clinical records, interpreting complex terminology, comparing documentation against evidence and generating coherent summaries from large volumes of information. For clinicians, operators, and administrative teams who…

Leer más
geekygadgetsai 10 Sep 2026

Mac Studio M5 Ultra vs Nvidia DGX Spark: Compared for Local AI

Selecting a system for local AI workloads involves evaluating performance, scalability and specific use case demands. RepoChad compares two prominent options: Apple’s Mac Studio M5 Ultra and Nvidia’s DGX Spark. The Mac Studio features 512 GB of unified memory and 1.2 TB/s bandwidth, making it well-suited for handling large-scale models, such as those with 700 […] The post Mac Studio M5 Ultra vs Nvidia DGX Spark: Compared for Local AI appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 10 Sep 2026

How to Work 10X Faster the Andrej Karpathy AI Workflow

Andrej Karpathy, a prominent figure in artificial intelligence, has adopted a thoughtful approach to incorporating AI into everyday tasks. One notable method involves using speech-to-text systems, such as Super Whisper, to interact with AI more naturally and efficiently. By dictating rather than typing, Karpathy minimizes friction in communication and enhances the flow of ideas. As […] The post How to Work 10X Faster the Andrej Karpathy AI Workflow appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 10 Sep 2026

Why GPT-6 Astra Fails the Open-Ended Goals Required for AGI

Artificial General Intelligence (AGI) remains a distant yet compelling goal in the field of artificial intelligence, promising systems capable of human-like adaptability and reasoning across diverse tasks. In their latest analysis, The Stack examines how models like GPT-6 Astra and Claude Fable 5.1 highlight both the strides and the persistent gaps in this pursuit. For […] The post Why GPT-6 Astra Fails the Open-Ended Goals Required for AGI appeared first on Geeky Gadgets.

Leer más
OpenAI 10 Sep 2026

Introducing ChatGPT for Financial Services

Introducing ChatGPT for Financial Services, combining built-in financial data and GPT-6 Astra for research, modeling, and client-ready materials.

Leer más
OpenAI 10 Sep 2026

Expanding AI access and cyber defense for federal, state, local, and tribal governments

OpenAI and GSA will offer eligible federal, state, local, and tribal governments $0 license fees, 50% off usage, and expanded cyber defense support.

Leer más
geekygadgetsai 10 Sep 2026

LangChain Connections Secures Managed Deep Agents in 3 Ways

Managing credentials is a critical aspect of deploying Managed Deep Agents (MDA) in environments that demand both security and scalability. LangChain addresses this need through its Connections system, which includes options like agent-owned secrets, user-owned OAuth via Managed Credential Providers (MCP), and user-owned OAuth with custom applications. For instance, agent-owned secrets allow a single API […] The post LangChain Connections Secures Managed Deep Agents in 3 Ways appeared first on Geeky Gadgets.

Leer más
OpenAI 10 Sep 2026

Build more natural voice experiences with GPT‑Live‑1 in the API

GPT‑Live‑1 brings natural, full-duplex voice conversations to the API, with stronger instruction following, custom voices, and telephony support.

Leer más
OpenAI 10 Sep 2026

Introducing the Agents API

Build and launch cloud agents with the Agents API, a managed service powered by the Codex harness for orchestration, long-running sessions, and tool use.

Leer más
HuggingFace 10 Sep 2026

Rebuilding AUTOMATIC1111 with Gradio Workflow

Leer más
HuggingFace 09 Sep 2026

IBM releases SOTA Granite Time Series PatchTST-FM-r2 model with commercial-friendly license

Leer más
geekygadgetsai 09 Sep 2026

DeepSeek V4.1 Flash Reaches 427 Tokens per Second in Tests

DeepSeek V4.1 Flash has emerged as a standout option for professionals seeking a balance of speed, affordability and versatility in AI-driven workflows. Developed by DeepSeek and available for testing until September 10, 2026, this model builds on the foundation of version 4 Flash, delivering enhanced performance while maintaining cost-efficiency. Notably, its processing speed reaches up […] The post DeepSeek V4.1 Flash Reaches 427 Tokens per Second in Tests appeared first on Geeky Gadgets.

Leer más
OpenAI 09 Sep 2026

The AI policy window is open. We need to act.

Chris Lehane argues that stronger AI capabilities require stronger safety evidence, shared standards, and durable policy action while the policy window remains open.

Leer más
geekygadgetsai 09 Sep 2026

Meet OpenAI Bell ChatGPT 7: the Model That Replaces GPT-6 Astra

OpenAI has introduced “Bell,” an advanced AI model that exceeds the capabilities of GPT-6 Astra and addresses significant challenges in computational science. A key achievement of “Bell” is its resolution of the Navier-Stokes Millennium Prize problem, a complex issue in fluid dynamics that had remained unsolved for decades. This accomplishment was made possible through the […] The post Meet OpenAI Bell ChatGPT 7: the Model That Replaces GPT-6 Astra appeared first on Geeky Gadgets.

Leer más
mitai 09 Sep 2026

The Download: OpenAI’s turning point for math and a battery record

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. What OpenAI’s latest controversy tells us about the future of math OpenAI says its agents have solved one of the most important open problems in mathematics. Under normal circumstances, that would…

Leer más
geekygadgetsai 09 Sep 2026

Leaked Meta Smart Glasses Lineup Reveals Four Upcoming Models

Meta’s upcoming smart glasses lineup has been revealed through recent leaks, showcasing four distinct models that cater to diverse user needs. Among these is the RW7004, a camera-free design that emphasizes privacy while still offering advanced audio and AI capabilities. The Smart Glasses Guy highlights how these glasses, developed in collaboration with Ray-Ban, combine style […] The post Leaked Meta Smart Glasses Lineup Reveals Four Upcoming Models appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 09 Sep 2026

ChatGPT Images 2.5 Turns Rough Sketches Into Final Art

ChatGPT Images 2.5 brings updates that improve the speed and quality of AI-generated visuals, offering new ways to approach creative projects. Paul Lipsky highlights the sketch-based image creation feature, which allows users to begin with a simple outline or concept sketch. The AI then refines this input into a polished, high-quality image. This feature supports […] The post ChatGPT Images 2.5 Turns Rough Sketches Into Final Art appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 09 Sep 2026

OpenAI Jalapeño Chip Delivers 9X Throughput in Benchmarks

OpenAI’s new Jalapeño chip, developed in partnership with Broadcom, marks the company’s first foray into custom hardware for AI inference tasks. This Application-Specific Integrated Circuit (ASIC) is engineered to deliver exceptional performance, boasting nearly nine times the throughput of competing hardware in specific benchmarks. Caleb Writes Code explores how this chip’s energy efficiency, processing 1,500 […] The post OpenAI Jalapeño Chip Delivers 9X Throughput in Benchmarks appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 09 Sep 2026

GLM 5.3 Flash Costs $0.24 per Task vs the $4 GLM 5.3 Model

Deciding between GLM 5.3 Flash and the full GLM 5.3 model involves weighing cost against performance and reliability. As highlighted by The Stack, Flash stands out for its cost efficiency, allowing users to complete tasks at just $0.24 each compared to the full model’s $4 per task. However, this affordability comes with trade-offs, such as […] The post GLM 5.3 Flash Costs $0.24 per Task vs the $4 GLM 5.3 Model appeared first on Geeky Gadgets.

Leer más
Gemini 09 Sep 2026

Get ready for the game with new football features in Search

Track live game feeds, explore detailed stats, and get custom fantasy recommendations directly in Search this season.

Leer más
Gemini 09 Sep 2026

Recreating a 70-year love story frame by frame

Discover how filmmakers and Google DeepMind used AI to recreate a couple's unrecorded past in the short film "Love, Rendered."

Leer más
geekygadgetsai 09 Sep 2026

Microsoft Copilot Can Now Build Microsoft PowerPoint Slides from OneNote

Microsoft Copilot in PowerPoint introduces a new way to create presentations by combining AI-driven features with user-friendly design options. As demonstrated by Mike Tholfsen, one standout capability is the ability to generate slides directly from existing documents, spreadsheets, or OneNote notebooks. This feature not only saves time but also ensures that the content aligns with […] The post Microsoft Copilot Can Now Build Microsoft PowerPoint Slides from OneNote appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 09 Sep 2026

Gemini Notebook Turns PDF Files Into Detailed Mind Maps for Research and More

Gemini Notebook, previously known as NotebookLM, is an AI-based platform created by Google to support research workflows for students, professionals and researchers. Andy Stapleton examines how it helps users organize and analyze various sources, such as PDFs and Google Drive files, while making sure references remain traceable for credibility. One notable feature highlighted is its […] The post Gemini Notebook Turns PDF Files Into Detailed Mind Maps for Research and More appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 09 Sep 2026

Google Pushes Gemini 4 Pro Release Date to October 2026

Google DeepMind’s Gemini 4 Pro has been delayed again, now projected for release in October 2026. The delay is attributed to pre-training challenges as the team works to refine the model to meet the benchmarks set by earlier versions. In the meantime, competitors like Enthropic and OpenAI are advancing their own offerings. Enthropic’s Fable 5.2, […] The post Google Pushes Gemini 4 Pro Release Date to October 2026 appeared first on Geeky Gadgets.

Leer más
HuggingFace 08 Sep 2026

Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic

Leer más
geekygadgetsai 08 Sep 2026

Local AI Hardware Requirements: More RAM vs Big GPU

Running advanced AI models like Qwen 3.8 Flash Next on local hardware presents unique challenges, particularly when deciding whether to prioritize more system memory (RAM) or a higher-capacity graphics card (GPU). For instance, Qwen 3.8 Flash Next, a 125-billion-parameter model, recommends 96 GB of RAM but lacks specific GPU memory requirements. This often forces users […] The post Local AI Hardware Requirements: More RAM vs Big GPU appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 08 Sep 2026

Build a Second Brain with ChatGPT 6 Astra to Automate Workflows

Nate Herk examines how ChatGPT 6 Astra can act as the backbone for building an AI Operating System (AIOS), a personalized system for managing knowledge and automating workflows. By using the “Four C’s” framework, Context, Connections, Capabilities and Cadence, you can design a scalable second brain tailored to your needs. For instance, incorporating Codeex during […] The post Build a Second Brain with ChatGPT 6 Astra to Automate Workflows appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 08 Sep 2026

Gemini 3.8 Flash Outperforms GPT 5.6 Sol in Coding Tests

Google’s latest AI model, Gemini 3.8 Flash, introduces significant advancements in reasoning, coding and multimodal functionality. Released in two variants, Gemini 3.8 Flash and Gemini 3.8 Flash Cyber, this model addresses diverse applications, including front-end design automation and real-time threat analysis. According to World of AI, one standout feature is its ability to integrate text, […] The post Gemini 3.8 Flash Outperforms GPT 5.6 Sol in Coding Tests appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 08 Sep 2026

Claude Needs Awwwards Layouts for Professional Web Design

Creating beautiful, functional websites with Claude involves more than just generating outputs, it requires thoughtful input and clear references. According to The AI Automators, Claude performs best when provided with specific design elements such as screenshots, defined color palettes, or typography samples. For example, referencing layouts from platforms like Awwwards or incorporating fonts from Google […] The post Claude Needs Awwwards Layouts for Professional Web Design appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

IFA Award-Winning viaim Rise AI Agent Earbuds Launch on Indiegogo September 8

In today’s fast-paced mobile workspace, we constantly move between virtual meetings, phone calls, and spontaneous brainstorming sessions, often struggling to retain every critical detail. While standard note-taking apps and recording tools catch basic audio, they leave us to manually sort transcripts and update task trackers ourselves. Fresh off the 2026 IFA Innovation Award for ‘Best […] The post IFA Award-Winning viaim Rise AI Agent Earbuds Launch on Indiegogo September 8 appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

Mac Studio Beats DGX Spark in local AI Tests with 1.2 TB/s Bandwidth

Choosing the right hardware for local AI workloads often comes down to balancing performance, usability and cost. In this deep dive, Kai examines the Mac Studio (M5 Ultra) and Nvidia DGX Spark, two machines designed for distinct use cases. For example, the Mac Studio stands out with its plug-and-play setup, allowing users to start working […] The post Mac Studio Beats DGX Spark in local AI Tests with 1.2 TB/s Bandwidth appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

Gemini Notebook Generates Built-in Spreadsheets for Business

Google DeepMind’s Gemini Notebook, formerly known as NotebookLM, has introduced AI-driven enhancements designed to support business data analysis and management. According to Universe of AI, the platform now enables tasks such as revenue analysis, customer behavior tracking and competitor research. A notable addition is its ability to generate spreadsheets directly within the interface, reducing reliance […] The post Gemini Notebook Generates Built-in Spreadsheets for Business appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

ChatGPT 6 Astra Defeats Fable 5.1 in 10 of 15 Test Scenarios

After 100 hours of testing across 15 distinct scenarios, Nate Herk presents a detailed comparison of GPT-6 Astra and Fable 5.1, focusing on their strengths and limitations in both technical and creative tasks. The overview reveals that Astra performed well in structured tasks like browser-based automation, offering cost efficiency and reliable outputs, while Fable excelled […] The post ChatGPT 6 Astra Defeats Fable 5.1 in 10 of 15 Test Scenarios appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

RISC-V XuanTie C950 CPU Reportedly Runs Qwen 3.8 Without a GPU

Alibaba’s recent announcement about its XuanTie C950 CPU has sparked interest in the AI hardware community. According to The Stack, this RISC-V-based processor reportedly runs the Qwen 3.8 model, a 27-billion-parameter AI system, without relying on a GPU. The CPU achieves a generation speed of 30 tokens per second, with the first token generated in […] The post RISC-V XuanTie C950 CPU Reportedly Runs Qwen 3.8 Without a GPU appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

Gemini 3.8 Flash Matches Opus 5 on Deep Sweep 1.1 Benchmark

Google’s release of Gemini 3.8 Flash and its specialized variant, Gemini 3.8 Flash Cyber, introduces updates aimed at balancing performance and accessibility. According to Universe of AI, features like High Mode, which achieves an intelligence index score of 59, highlight the model’s focus on practical use cases. While not designed to compete with high-end systems […] The post Gemini 3.8 Flash Matches Opus 5 on Deep Sweep 1.1 Benchmark appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 07 Sep 2026

ChatGPT 6 Astra Builds a Playable FPS Game in 30 Minutes

OpenAI’s ChatGPT 6 Astra represents a significant advancement in artificial intelligence, offering new ways to approach complex challenges across various fields. According to World of AI, one standout example is its application in game development, where it can generate fully functional games in a fraction of the time traditionally required. In one instance, Astra created […] The post ChatGPT 6 Astra Builds a Playable FPS Game in 30 Minutes appeared first on Geeky Gadgets.

Leer más
grokwiki 07 Sep 2026

Muskxxx06: Update 2026 models and DoD GenAI.mil availability; add sourced legal notes. Grokipedia - Note Lawfare / Verge reporting that edits appear stalled after 24 Apr 2026.

Leer más
geekygadgetsai 05 Sep 2026

Claude Fable 5.1 Adds 75% Repeat Discount but Higher Cache Fees

Claude Fable 5.1 brings a revised pricing model that appears promising for workflows involving repetitive tasks, but its benefits come with notable trade-offs. As detailed by The Stack, the model introduces a discounted rate of $0.25 per million tokens for repeated inputs, provided they are sent within a five-minute window and remain unchanged. However, the […] The post Claude Fable 5.1 Adds 75% Repeat Discount but Higher Cache Fees appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 05 Sep 2026

Everthine A1 brings a 3D AI companion to IFA 2026

Everthine is showing its A1 desktop AI companion at IFA 2026 in Berlin, giving European visitors their first chance to see the dedicated device in action. The showcase runs September 4-8 at Messe Berlin and comes ahead of the product’s planned release later this year. A1 places a real-time 3D character on an 8.01-inch curved […] The post Everthine A1 brings a 3D AI companion to IFA 2026 appeared first on Geeky Gadgets.

Leer más
githubcopilot 04 Sep 2026

Project HydraFusion: Frontier quality via multi-model orchestration

In controlled offline evaluations, HydraFusion’s selective coding workflows matched or exceeded the evaluated Opus 5 baseline while reducing estimated workflow cost. Now available as a research preview in GitHub Copilot. The post Project HydraFusion: Frontier quality via multi-model orchestration appeared first on The GitHub Blog.

Leer más
geekygadgetsai 04 Sep 2026

Build an AI Email Agent with OpenAI ChatGPT and the Gmail API: Beginners Guide

Managing a high-volume inbox can quickly become a time-consuming challenge, but building an AI email agent offers a structured way to automate this task. In this step-by-step guide, Corbin explains how to create a system that integrates with platforms like Gmail or Outlook, uses AI models such as OpenAI’s GPT and incorporates customizable rules to […] The post Build an AI Email Agent with OpenAI ChatGPT and the Gmail API: Beginners Guide appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 04 Sep 2026

Google Gemini 3.8 Flash Costs $0.75 per Million Input Tokens

Google’s latest release, Gemini 3.8 Flash, marks a significant addition to the AI landscape, offering specialized capabilities tailored to specific industries. As highlighted by Matthew Berman, the model excels in targeted benchmarks such as the Harvey Legal Benchmark, where it achieves a leading score of 61.4%, and Terminal Bench 2.1, with an impressive 89.4% for […] The post Google Gemini 3.8 Flash Costs $0.75 per Million Input Tokens appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 04 Sep 2026

Claude Fable 5.1 Launches with Strict Safety but Less Flexibility

Anthropic’s release of Claude Fable 5.1 introduces a model aimed at enhancing safety and compliance, making it particularly relevant for sectors like finance and healthcare where managing risk is essential. According to Universe of AI, this version builds on Mythos 5.1 by adding stricter safeguards, such as limiting cybersecurity-related functionalities, to ensure controlled applications. While […] The post Claude Fable 5.1 Launches with Strict Safety but Less Flexibility appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 04 Sep 2026

Rokid Air Glasses Style 10-Minute Video Beats Ray-Ban Meta Gen 2

Three months ago, Tech with Spencer shared initial impressions of the Rokid Ai Glasses Style, highlighting its lightweight design and promising AI capabilities. Now, after extended use, the glasses have shown notable growth, particularly through regular software updates that have refined their performance. Features like real-time translation and customizable controls have become more reliable, addressing […] The post Rokid Air Glasses Style 10-Minute Video Beats Ray-Ban Meta Gen 2 appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 03 Sep 2026

ChatGPT 6 Officially Launches as Astra: Beats GPT-5.6 Sol with 47% Faster Speed

OpenAI has officially launched the highly anticipated ChatGPT 6 Astra, an AI model designed for complex professional applications. This release brings a 47% improvement in task completion speed compared to its predecessor while reducing token usage. Notable capabilities include secure code reviews and scientific data analysis, addressing needs in fields where accuracy and reliability are […] The post ChatGPT 6 Officially Launches as Astra: Beats GPT-5.6 Sol with 47% Faster Speed appeared first on Geeky Gadgets.

Leer más
Llama 03 Sep 2026

ZGateway: Learnings from Putting a Proxy in Front of ZippyDB

We’re introducing ZGateway, the proxy we are using to unify traffic through ZippyDB, Meta’s most widely-used key value store. As a bonus, it also enables admission control, load balancing, cross-region resilience, and richer operations. ZippyDB is the most widely used key value store at Meta, backing product metadata, counters, and configuration, and can serve billions [...] Read More... The post ZGateway: Learnings from Putting a Proxy in Front of ZippyDB appeared first on Engineering at Meta.

Leer más
githubcopilot 03 Sep 2026

GitHub Copilot app for Beginners: Run several agents at once

Learn how to run parallel agents in the GitHub Copilot app, and experience the moment it stops feeling scary and starts feeling powerful. The post GitHub Copilot app for Beginners: Run several agents at once appeared first on The GitHub Blog.

Leer más
HuggingFace 03 Sep 2026

NeoMME: an efficient Multimodal-native and Multilingual Encoder

Leer más
geekygadgetsai 03 Sep 2026

RayNeo iO Smart Glasses Pack AI Into a 34g Titanium Frame

The RayNeo iO smart glasses bring advanced functionality into everyday scenarios, blending practicality with thoughtful design. As noted by Chigz Tech Reviews, these glasses feature a lightweight 34-gram frame and a transparent waveguide display, making them comfortable for extended use. Key highlights include real-time conversation translation and voice-controlled reminders, which cater to both professional and […] The post RayNeo iO Smart Glasses Pack AI Into a 34g Titanium Frame appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 03 Sep 2026

Fable 5.1 Upgrades AI Web Design with Layered Scrolling

Fable 5.1 brings a refined approach to AI-driven website design, addressing longstanding challenges in functionality and user experience. According to Nate Herk, this update introduces features like layered scrolling effects and touch-friendly navigation, alongside responsive grids that cater to mobile-first design priorities. These additions not only enhance visual coherence but also ensure websites perform reliably […] The post Fable 5.1 Upgrades AI Web Design with Layered Scrolling appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 03 Sep 2026

Gemini Notebook Replaces Google NotebookLM for Data Analytics

Google has rebranded NotebookLM as Gemini Notebook, introducing features designed to support more dynamic and structured workflows. One notable update is the Agentic Research capability, which helps users organize fragmented ideas and incomplete datasets into coherent outputs. According to Universe of AI, this feature is particularly effective for tasks like market trend analysis, where it […] The post Gemini Notebook Replaces Google NotebookLM for Data Analytics appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 03 Sep 2026

New Google Gemini 3.8 Flash Matches Opus 5 in Deep Sweep 1.1 Benchmark

Google’s recent release of Gemini 3.8 Flash has introduced a model that balances advanced performance with cost efficiency, catering to both technical and enterprise needs. As highlighted by Prompt Engineering, this model excels in specific benchmarks like Deep Sweep 1.1, demonstrating its ability to handle complex computational tasks. However, it comes with trade-offs, such as […] The post New Google Gemini 3.8 Flash Matches Opus 5 in Deep Sweep 1.1 Benchmark appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 03 Sep 2026

Claude AI Adds Invisible Text Watermarks to Everything for European AI Act

Claude AI’s introduction of an invisible watermark in all its generated text represents a significant development in making sure transparency and accountability in AI-generated content. According to Ryan & Matt Data Science, this watermark uses subtle, pattern-based algorithms embedded during text generation, making it undetectable to human readers but identifiable through Claude-specific detection systems. While […] The post Claude AI Adds Invisible Text Watermarks to Everything for European AI Act appeared first on Geeky Gadgets.

Leer más
HuggingFace 03 Sep 2026

Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps

Leer más
HuggingFace 03 Sep 2026

Give Your Coding Agents a Memory You Own

Leer más
HuggingFace 03 Sep 2026

Training a coding model to paint watercolours with TRL and OpenEnv

Leer más
githubcopilot 02 Sep 2026

How we make AI coding more cost efficient without sacrificing task quality

Why shorter outputs can cost more, and how GitHub Copilot reduces wasted work across the complete coding task. The post How we make AI coding more cost efficient without sacrificing task quality appeared first on The GitHub Blog.

Leer más
geekygadgetsai 02 Sep 2026

VoxCPM2 Needs Just 8GB VRAM to Clone and Generate Any Voice Locally

VoxCPM2, developed by OpenBMBB, introduces a self-hosted approach to text-to-speech (TTS) technology, combining voice generation, cloning and multilingual synthesis into one cohesive system. Unlike cloud-based alternatives, this local AI model prioritizes data privacy and operational control, making it particularly appealing for industries with strict confidentiality requirements. For example, it allows users to create entirely new […] The post VoxCPM2 Needs Just 8GB VRAM to Clone and Generate Any Voice Locally appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 02 Sep 2026

Expected Sept 3 ChatGPT 6 Release Carries Hidden Reasoning Risks

OpenAI’s latest AI model, Astra possibly ChatGPT 6, has introduced a new concept known as “recurrent depth”, which enables the system to refine its understanding of input data through iterative processing. Unlike traditional models that process information in a single pass, GPT-6 Astra’s architecture revisits data multiple times, allowing it to generate more nuanced and […] The post Expected Sept 3 ChatGPT 6 Release Carries Hidden Reasoning Risks appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 02 Sep 2026

Claude Fable 5.1 Launches with a 25% Drop in Token Consumption

Fable 5.1 has officially arrived, bringing with it a host of updates aimed at refining AI-driven workflows. As highlighted by Nate Herk, one of the standout improvements is a 25% reduction in token consumption for standard workflows, which directly enhances both processing speed and resource efficiency. This makes it particularly valuable for tasks involving advanced […] The post Claude Fable 5.1 Launches with a 25% Drop in Token Consumption appeared first on Geeky Gadgets.

Leer más
Llama 02 Sep 2026

An Organizational Second Brain: Building an AI That Learns From Experts

We’ve built an AI agent that acts as a secondary expert for a given domain, making deep specialist knowledge readily available and preserved for anyone in an organization to access, share, and build upon. This is not a typical domain-specific agent. Its novelty comes from integrating two layers:  A structured, auditable knowledge architecture separates what [...] Read More... The post An Organizational Second Brain: Building an AI That Learns From Experts appeared first on Engineering at Meta.

Leer más
geekygadgetsai 02 Sep 2026

New Anthropic Claude Fable 5.1 Tops Terminal Bench Science Tests

Anthropic’s latest AI models, Fable 5.1 and Mythos 5.1, aim to push the boundaries of AI performance while addressing key enterprise concerns. Fable 5.1 is designed for precision in complex reasoning and agentic tasks, excelling in benchmarks like Terminal Bench Science, though it comes with a steep price tag. Mythos 5.1, on the other hand, […] The post New Anthropic Claude Fable 5.1 Tops Terminal Bench Science Tests appeared first on Geeky Gadgets.

Leer más
Gemini 02 Sep 2026

Proactive cyber defense for governments and enterprises

The Fairwind Program is a limited access program for governments and trusted partners to use our cyber defense tools.

Leer más
geekygadgetsai 02 Sep 2026

Make a Tiny Pocket ESP32 Al Assistant to Manage Thermostats and Lights

Building a pocket-sized AI assistant might seem like a daunting challenge, but Huy Vector demonstrates how it’s possible to create a device that blends functionality with personality. Using the ESP by We Vector as a foundation, this project showcases how a compact design can deliver features like voice-controlled smart home integration, real-time weather updates and […] The post Make a Tiny Pocket ESP32 Al Assistant to Manage Thermostats and Lights appeared first on Geeky Gadgets.

Leer más
microsoftblog 02 Sep 2026

The yield imperative: Turning AI infrastructure into useful intelligence

As we enter the next era, what will be the defining measure of our progress? Every industry has a word that shapes how it thinks. For pilots, it’s safety. For insurers, it’s risk. For the semiconductor industry, it’s yield. Yield does not ask how elegant the solution is, how many years it took or what... The post The yield imperative: Turning AI infrastructure into useful intelligence appeared first on The Official Microsoft Blog.

Leer más
HuggingFace 01 Sep 2026

BenchMIRT: What are LLM benchmarks actually measuring?

Leer más
Gemini 01 Sep 2026

The latest AI news we announced in August 2026

Here are Google’s latest AI updates from August 2026

Leer más
geekygadgetsai 01 Sep 2026

Tencent HY4 Brings 1 Million Token Context to Apache 2.0

Tencent’s HY4 is an open-weight AI model that combines efficiency with advanced computational design, as highlighted by World of AI. It features 770 billion parameters, with only 49 billion active at any time, thanks to its Mixture of Experts (MoE) architecture. This approach ensures optimized resource use while maintaining high performance. Additionally, its 1 million […] The post Tencent HY4 Brings 1 Million Token Context to Apache 2.0 appeared first on Geeky Gadgets.

Leer más
geekygadgetsai 01 Sep 2026

ChatGPT 6 Astra Reportedly Generates 3D Games from a Single Prompt

OpenAI’s ChatGPT 6 Astra, Anthropic’s Claude Opus 5.1 and Tencent’s HY4 represent distinct approaches to advancing artificial intelligence, each with its own strengths and challenges. In this hands-on review, World of AI explores how GPT-6 Astra’s ability to generate complex outputs like interactive 3D games and user interfaces from a single prompt positions it as […] The post ChatGPT 6 Astra Reportedly Generates 3D Games from a Single Prompt appeared first on Geeky Gadgets.

Leer más
Gemini 01 Sep 2026

Try Google Pics: Easy image creation and editing in Google Workspace

Built on our latest Nano Banana model, Google Pics — our image creation and editing tool — is now available.

Leer más
HuggingFace 01 Sep 2026

Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI

Leer más
HuggingFace 28 Aug 2026

The Open ASR Leaderboard Adds Its First Global South Language

Leer más
Gemini 27 Aug 2026

3 new ways to plan and book travel in Search

Book hotels and track airfares, plus view miles and rewards with AI Mode in Google Search.

Leer más
githubcopilot 26 Aug 2026

GitHub Copilot app for Beginners: Automate Dependabot pull request triage

Managing library updates can be tedious at times. Learn how the GitHub Copilot app can handle this type of repetitive task. The post GitHub Copilot app for Beginners: Automate Dependabot pull request triage appeared first on The GitHub Blog.

Leer más
Gemini 25 Aug 2026

5 ways to upgrade your home decor with Google Search

Learn how to use Google Search tools to find home decor inspiration, shop for furniture, and tackle DIY projects.

Leer más
Llama 24 Aug 2026

MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet

Training and serving frontier AI models depends on fast, reliable networks that move data between GPUs without wasting compute cycles.  To meet this challenge at scale, Meta designed MetaRoCE – a clean-sheet RDMA transport protocol purpose-built for AI workloads on commodity Ethernet.  We're releasing the MetaRoCE specification, a reference software implementation and a compliance test [...] Read More... The post MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet appeared first on Engineering at Meta.

Leer más
Llama 24 Aug 2026

MTIA 300: Meta’s First Training Chip with Built-in NICs and Communication-Offloading Engines

MTIA 300 is the first of Meta’s family of in-house training and inference accelerators optimized for training ranking and recommendation models. We’re sharing how MTIA 300’s built-in NIC chiplets allow it to meet the communication needs associated with training recommendation models with superior performance over general-purpose GPUs. By co-designing MTIA’s communication library, HCCL, alongside the [...] Read More... The post MTIA 300: Meta's First Training Chip with Built-in NICs and Communication-Offloading Engines appeared first on Engineering at Meta.

Leer más
Gemini 19 Aug 2026

5 new ways to level up your learning with Search

Here’s how you can use Google Search tools to study for classes and standardized tests.

Leer más
Gemini 17 Aug 2026

Get closer to the game with Gemini and Pixel

Google Gemini and Pixel partner with five global football clubs to elevate the fan matchday experience through AI and Smartphone Technology.

Leer más
Llama 12 Aug 2026

How We’re Building Scam Alert on WhatsApp With End-to-End Encryption and Verifiability Guarantees

WhatsApp is committed to helping people stay safe while protecting the privacy of their messages. As scam tactics evolve — from impersonation to social engineering to AI-generated lures — we're always evolving as well, so that our protections stay ahead of scammers while protecting people’s personal messages with end-to-end encryption. Today, we're sharing an early [...] Read More... The post How We’re Building Scam Alert on WhatsApp With End-to-End Encryption and Verifiability Guarantees appeared first on Engineering at Meta.

Leer más
Llama 05 Aug 2026

From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking

Every day, Meta’s recommendation platforms handle billions of user interactions, generating rich temporal signals that capture individual preferences and intent across products, ads, and content. In our 2024 post on sequence learning for ads recommendations, we showed how modeling the order and timing of user actions (rather than relying on static, manually engineered sparse features) [...] Read More... The post From User Sequences to Scaling Laws: A Multi-Stage Architecture for Meta’s Ads Ranking appeared first on Engineering at Meta.

Leer más
Llama 03 Aug 2026

GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model

Meta's Generative Ads Recommendation Model (GEM), the foundation model behind ads recommendations across Instagram and Facebook, now trains at LLM scale on several thousand of the latest-generation GPUs. This post goes into the details on how we achieved: doubling end-to-end (E2E) training efficiency to 20–25% Model FLOPs Utilization (MFU) while scaling training FLOPs 4x in [...] Read More... The post GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model appeared first on Engineering at Meta.

Leer más
microsoftblog 28 Jul 2026

Looking back on Microsoft’s FY26: From AI experimentation to Frontier Transformation

Throughout this past fiscal year, customers across every industry and segment moved from AI experimentation to deploying AI for real-world business outcomes. They unlocked innovation and created new opportunities for growth. We saw the emergence of Frontier Firms as they moved beyond efficiency gains to focus on human ambition and embed AI at the core... The post Looking back on Microsoft’s FY26: From AI experimentation to Frontier Transformation appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 27 Jul 2026

Rethinking security for the age of AI

Editor's note: Updates with additional details on the model's crash score. Why security needs a new cyber stack — Introducing Project Perception The physics of cybersecurity are changing. Autonomous systems can now reason, adapt and operate continuously. At the same time, the cost of offense is falling, while the volume, velocity and complexity of what... The post Rethinking security for the age of AI appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 22 Jul 2026

Powering America’s Genesis Mission: Microsoft’s commitment to scientific discovery

Today, we’re excited to share a long-term commitment to the Department of Energy’s (DOE) Genesis Mission, backed by a $60 million investment designed to accelerate AI for science and the breakthroughs it can deliver for the country. This deepened commitment includes Microsoft’s new Scientific Partnership Advancing Research & Knowledge coordination hub and program office, otherwise... The post Powering America's Genesis Mission: Microsoft's commitment to scientific discovery appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 20 Jul 2026

Microsoft expands Azure AI and HPC infrastructure with AMD

AI workloads are scaling faster than any single infrastructure approach can support — with more models, new agent-driven workloads and surging compute demand driving the need for greater specialization across the stack. To meet this need, Microsoft continues to evolve Azure’s infrastructure, including expanding its AI fleet with AMD’s most advanced AI and high-performance computing... The post Microsoft expands Azure AI and HPC infrastructure with AMD appeared first on The Official Microsoft Blog.

Leer más
Llama 15 Jul 2026

Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization

Hierarchical Interest Representation is a research area for Meta Ads. We’re exploring an upstream representation layer over the universe of Ads entities – users, advertisers, products, services – learning unified embeddings that connect users' inferred interests with the breadth of what advertisers offer in their deep funnel ads. The innovations in Hierarchical Interest Representation are [...] Read More... The post Exploring Hierarchical Interest Representation For Meta Ads Deep Funnel Optimization appeared first on Engineering at Meta.

Leer más
Llama 13 Jul 2026

Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler

TL; DR At Meta's scale, a few milliseconds of latency degradation can have a significant negative impact on ads performance.  When a Linux kernel upgrade risked regressing latency across Meta's ad serving fleet, we turned to sched_ext — the upstream, BPF-based extensible scheduling framework — to build a scheduling policy customized to the Ads delivery [...] Read More... The post Modernizing the Meta Ads Service With an Open-Source Kernel Scheduler appeared first on Engineering at Meta.

Leer más
microsoftblog 06 Jul 2026

The latest in our company transformation

Amy Coleman, EVP and Chief People Officer, shared the following communication with employees today. When I stepped into this role, I promised to communicate more openly with you and share the “why” behind our decisions. Today we are eliminating around 4,800 roles, about 2.1% of our global workforce, as we focus our people, investments, and... The post The latest in our company transformation appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 02 Jul 2026

Microsoft Frontier Company: AI engineering that amplifies and protects your intelligence

The pace of AI adoption is moving incredibly fast. Customers have moved well beyond experimentation and understand the importance of adopting AI to transform their business. They are now concentrating on delivering measurable business outcomes and demonstrating a return on their AI investments, while ensuring their intelligence is amplified and their IP is protected. Today... The post Microsoft Frontier Company: AI engineering that amplifies and protects your intelligence appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 24 Jun 2026

Inside Microsoft’s two-decade push to cut water intensity while scaling for growth

As demand for cloud and AI services continues to grow, datacenters are becoming more essential than ever. Communities also want to better understand how this infrastructure affects local resources, particularly water. At Microsoft, water stewardship has been a priority since our first datacenter builds in the early 2000s and remains a core part of our... The post Inside Microsoft’s two-decade push to cut water intensity while scaling for growth appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 23 Jun 2026

Rethinking cloud operations with agentic observability

Cloud operations are entering a new era as AI-driven and autonomous agents become a larger part of modern software systems. As software becomes increasingly agentic, the challenge is no longer just managing greater scale and complexity. Operators must also contend with systems that evolve faster, act more autonomously and interact across an expanding network of... The post Rethinking cloud operations with agentic observability appeared first on The Official Microsoft Blog.

Leer más
microsoftblog 22 Jun 2026

Powering the next wave of AI: Expanding capacity with our new datacenter in Pecos

Today, Microsoft is announcing one of the largest single capacity additions in our history. In Pecos, Texas, we will build a new datacenter campus, expanding our global datacenter capacity by approximately 2 gigawatts (GW) to meet strong and sustained customer demand for AI and cloud services across industries and regions. Beyond the technology, this is... The post Powering the next wave of AI: Expanding capacity with our new datacenter in Pecos appeared first on The Official Microsoft Blog.

Leer más
Claude 29 Sep 2025

Effective context engineering for AI agents

Effective context engineering for AI agents

Leer más
Claude 17 Sep 2025

A postmortem of three recent issues

A postmortem of three recent issues

Leer más
Claude 11 Sep 2025

Writing effective tools for agents — with agents

Writing effective tools for agents — with agents

Leer más
Claude 26 Jun 2025

Desktop Extensions: One-click MCP server installation for Claude Desktop

Desktop Extensions: One-click MCP server installation for Claude Desktop

Leer más
Claude 13 Jun 2025

How we built our multi-agent research system

How we built our multi-agent research system

Leer más
Claude 18 Apr 2025

Claude Code: Best practices for agentic coding

Claude Code: Best practices for agentic coding

Leer más
Claude 20 Mar 2025

The "think" tool: Enabling Claude to stop and think in complex tool use situations

The "think" tool: Enabling Claude to stop and think in complex tool use situations

Leer más
Claude 06 Jan 2025

Raising the bar on SWE-bench Verified with Claude 3.5 Sonnet

Raising the bar on SWE-bench Verified with Claude 3.5 Sonnet

Leer más
Claude 19 Dec 2024

Building effective agents

Building effective agents

Leer más
Claude 19 Sep 2024

Introducing Contextual Retrieval

Introducing Contextual Retrieval

Leer más