MarkTechPost
AI news source. Articles are auto-selected and adapted by Hamidun News editors.
Latest publications

Meta AI Presented Brain2Qwerty v2 — Non-invasive MEG Text Decoder with 61% Accuracy
Meta AI released Brain2Qwerty v2 — a non-invasive MEG-based pipeline that decodes typed sentences from brain signals with 61% word-level accuracy and published the training code.

Sakana AI Released Sakana Marlin — an Agent for 100-Page Reports with Slides
Sakana AI Introduced Its First Commercial Product — the Corporate Agent Marlin, Which Works Autonomously for Up to Eight Hours and Generates Reports Up to 100 Pages with Slides

Baidu Develops CUP — Set of Python Tools for Reliable Backend Pipelines
MarkTechPost analyzed CUP — a set of Python utilities from Baidu for logging, caching, thread pools and Linux resource monitoring in one package.

Google Research Releases TabFM — Foundation Model for Tabular Data with Zero-Shot
Google Research presented TabFM — a foundation model for tables that performs zero-shot classification and regression without training for each dataset.

Linq Launched iMessage Apps — Payments, Tickets and Games for AI Agents Directly in Chats
Linq introduced iMessage Apps — interactive imessage_app cards for payments, ticket purchases, flight bookings and games within messaging.

Claude Sonnet 5 vs Opus 4.8: agentic coding comparison and API pricing from Anthropic
MarkTechPost compared Claude Sonnet 5, Sonnet 4.6, and Opus 4.8 from Anthropic: the new model approaches the flagship in agentic coding at lower Sonnet token prices.

Anthropic resumed Claude Fable 5 operations after US export restrictions were lifted
Anthropic on July 1, 2026 restored access to Claude Fable 5 and introduced a classifier that blocks jailbreak techniques.

OpenAI Released GPT-Realtime-2.1 and Mini Version for Voice Agents in API
OpenAI added two models for voice agents to its API—GPT-Realtime-2.1 and GPT-Realtime-2.1-mini—reducing p95 latency by at least 25% through caching.

Tencent released Hy3—an open MoE model with 295 billion parameters and 256K context
Tencent presented Hy3—a model with 295 billion parameters, 21 billion active per token, Apache 2.0 license and a 78.0 score on SWE-Bench Verified.

MarkTechPost showed how to build an AI co-researcher QSAR for finding EGFR inhibitors
A MarkTechPost tutorial teaches how to build an autonomous AI co-researcher using Random Forest QSAR — from ChEMBL and RDKit data to ready-made candidate molecules against EGFR C797S.

Liquid AI released LFM2.5-350M: efficient 350M parameter model with scaled RL
Liquid AI proved that AI density is more important than parameter count: small model can be trained better through reinforced learning.

Parallax: improved local linear attention with covariance correction
Researchers from MarkTechPost presented Parallax — a new variation of the attention mechanism for small LLMs with improved quality and speed.

MiniMax M3: New Model with 1M Tokens, Multimodality, and Computer Use
The company introduced the MiniMax Sparse Attention architecture supporting one million tokens in context and work with video, images, and computer control.

Autonomous AI agents run 50x longer than search: Harvard and Perplexity research
New research from Harvard and Perplexity shows AI agents perform 26 minutes of independent work per session, compared to 33 seconds for search assistants.

Memory OS: Open 6-Layer Memory Stack for Hermes Agent
New Memory OS project adds persistent local memory to Hermes Agent through six architecture layers, managed memory search, and embedded wiki.

Google launches Gemini 3.5 Live Translate: speech-to-speech model for 70+ languages
Gemini 3.5 Live Translate translates speech in real time, synthesizing audio with a lag of several seconds. The model is available in Gemini Live API, Google Meet, and Google Translate app.

Nous Research presented Hermes Agent profile builder
New Hermes Agent interface replaces multi-step CLI configuration — users now assemble agent profiles in a single dashboard flow.

SpaceXAI Released Grok 4.5 — Opus-Class Code Model
SpaceXAI presented Grok 4.5, Opus-class model trained with Cursor for coding and agentic tasks, priced at $2/$6 per million tokens.

Ant Group Presented LingBot-VLA 2.0 — Open Model for Controlling Robots of Any Design
Its 6-billion parameter model trained on 60 thousand hours of robotics video and can control a dozen types of manipulators through a unified action space.

NVIDIA Nemotron-Labs-3-Puzzle-75B: 37% model compression doubled server speed
NVIDIA released a compressed version of the Nemotron-3-Super LLM — Nemotron-Labs-3-Puzzle-75B. Thanks to the Puzzle method (alternating compression and distillation), the model became 37% smaller and doubled server throu

Google Research released SensorFM — a model for 35 health forecasting tasks
Google Research trained the foundational SensorFM model on 1 trillion minutes of data from 5 million Fitbit and Pixel Watch users to predict health outcomes.

Robbyant releases LingBot-World-Infinity, a causal model for interactive worlds
Robbyant unveiled LingBot-World-Infinity, a 14B video model that works as an interactive simulator, generating 720p video at 60 frames per second.

Datalab Lift vs Competitors: How a 9B Document Extractor Works with JSON Schema
Datalab compared its 9B Lift tool with NuExtract3, LlamaExtract, Marker, and Docling — and analyzed when the schema-first approach wins in PDF data extraction.

Liquid AI Released Antidoom: FTPO Method Reduces Hangs in Reasoning Models
Liquid AI open-sourced Antidoom — a tool that identifies the token trigger of infinite loops and retrains only it, reducing doom-loop frequency from 22.9% to 1%.

Google AI Studio added GitHub repository import to Build development mode
Google AI Studio updated Build mode — any GitHub repository can now be imported directly and immediately turned into an editable, deployable app.

NVIDIA released Audex-30B — a unified audio-text MoE model based on Nemotron-Cascade-2
NVIDIA combined speech recognition, translation, speech synthesis, and audio generation in a single 30-billion-parameter MoE model while preserving Nemotron-Cascade-2’s text capabilities.

Ant Group open-sources LingBot-Vision: a 1B model for robot spatial perception
Ant Group's Robbyant has released LingBot-Vision, a 1B ViT model trained to detect object boundaries without labels and outperform larger counterparts on spatial perception tasks.

AI for PDF Document Processing: 2026 Tools Guide
Overview of open-source tools for converting corporate PDFs and scans into structured JSON — essential for LLMs and AI agents to work with real corporate data.

Meituan Releases LongCat-2.0: Open MoE Model with 1.6 Trillion Parameters and 1 Million Token Context
Chinese company Meituan unveiled LongCat-2.0 — a massive MoE model with native one-million-token context window trained on domestic ASICs.

Synthetic Sciences Releases OpenScience — Open AI Workbench for ML, Biology, Physics and Chemistry
OpenScience by Synthetic Sciences — open-source workbench under Apache 2.0 with 250+ skills, compatible with any AI model for ML, biology, physics and chemistry tasks.