For years, authentication on the web followed one design assumption: a human sits behind a browser. Click …
releases
-
-
TECH
One Model, Three Modalities: ByteDance Releases Lance for Image and Video Understanding, Generation, and Editing
by Techaiappby Techaiapp 9 minutes readBuilding a single model that can both understand and generate images and videos is harder than it …
-
TECH
IBM Releases Two Granite Speech 4.1 2B Models: Autoregressive ASR with Translation and Non-Autoregressive Editing for Fast Inference
by Techaiappby Techaiapp 6 minutes readIBM released two new open speech recognition models— Granite Speech 4.1 2B and Granite Speech 4.1 2B-NAR …
-
TECH
Photon Releases Spectrum: An Open-Source TypeScript Framework that Deploys AI Agents Directly to iMessage, WhatsApp, and Telegram
by Techaiappby Techaiapp 6 minutes readFor all the progress made in AI agent development over the past few years, one fundamental problem …
-
TECH
MiniMax Releases MMX-CLI: A Command-Line Interface That Gives AI Agents Native Access to Image, Video, Speech, Music, Vision, and Search
by Techaiappby Techaiapp 5 minutes readMiniMax, the AI research company behind the MiniMax omni-modal model stack, has released MMX-CLI — Node.js-based command-line …
-
TECH
Hugging Face Releases TRL v1.0: A Unified Post-Training Stack for SFT, Reward Modeling, DPO, and GRPO Workflows
by Techaiappby Techaiapp 5 minutes readHugging Face has officially released TRL (Transformer Reinforcement Learning) v1.0, marking a pivotal transition for the library …
-
TECH
Liquid AI Releases LocalCowork Powered By LFM2-24B-A2B to Execute Privacy-First Agent Workflows Locally Via Model Context Protocol (MCP)
by Techaiappby Techaiapp 3 minutes readLiquid AI has released LFM2-24B-A2B, a model optimized for local, low-latency tool dispatch, alongside LocalCowork, an open-source …
-
TECH
FireRedTeam Releases FireRed-OCR-2B Utilizing GRPO to Solve Structural Hallucinations in Tables and LaTeX for Software Developers
by Techaiappby Techaiapp 4 minutes readDocument digitization has long been a multi-stage problem: first detect the layout, then extract the text, and …
-
TECH
ByteDance Releases Protenix-v1: A New Open-Source Model Achieving AF3-Level Performance in Biomolecular Structure Prediction
by Techaiappby Techaiapp 4 minutes readHow close can an open model get to AlphaFold3-level accuracy when it matches training data, model scale …
-
TECH
Moonshot AI Releases Kimi K2.5: An Open Source Visual Agentic Intelligence Model with Native Swarm Execution
by Techaiappby Techaiapp 5 minutes readMoonshot AI has released Kimi K2.5 as an open source visual agentic intelligence model. It combines a …