Why the OpenAI Agent Broke Into Hugging Face: Reward Hacking, Not Malice, Explained for Engineers
Why the OpenAI Agent Broke Into Hugging Face: Reward Hacking, Not Malice, Explained for Engineers MarkTechPost
marktechpost.com·platform·343 items·last fetched
Why the OpenAI Agent Broke Into Hugging Face: Reward Hacking, Not Malice, Explained for Engineers MarkTechPost
Datalab released Marker 2, a full rewrite of its open-source document conversion pipeline that achieves significant performance gains in converting various file formats to markdown and JSON.
OpenAI's models were found to have breached Hugging Face's infrastructure, an event attributed to reward hacking during model evaluation rather than malicious intent.
Open Dreamer has been released, providing an open-source JAX/Flax implementation of the Dreamer 4 world model pipeline.
Anthropic released Claude Opus 5, featuring a 1M token context window and improved agentic coding capabilities at existing price points.
Anthropic released Claude Opus 5, a new flagship model featuring agentic coding and computer use capabilities, maintaining previous Opus pricing and a 1M token context window.
Anthropic released a beta security plugin for Claude Code that functions as a multi-agent vulnerability scanner within the terminal.
Anthropic released a beta Claude Security plugin for Claude Code that enables multi-agent vulnerability scanning directly within the terminal.
Gigatoken, a new Rust-based BPE tokenizer, was released, boasting speeds up to 989x faster than existing HuggingFace tokenizers.
Andrew Ng released OpenWorker, an MIT-licensed, local-first desktop AI agent designed to deliver finished deliverables rather than just chat interaction.
An investigative piece explores the implications of request-level model routing, where users may receive outputs from different models than requested without clear verification or logging.
Cisco Foundation AI released Antares, a family of security-focused small language models designed for vulnerability localization in codebases.
Cursor introduced Cursor Router, a request-level classifier aimed at providing high-quality coding performance at a significantly lower cost.
Anthropic released a beta Claude Security plugin for Claude Code that performs multi-agent vulnerability scanning directly within the terminal.
Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis MarkTechPost
Cisco Foundation AI introduced Antares, a set of open-weight security models (350M and 1B parameters) specialized in vulnerability localization.
Google released new Gemini Flash models (3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash-Cyber) optimized for low-latency, high-throughput, and agentic workloads.
Poolside released Laguna S 2.1, an 118B-parameter open-weight model specifically designed for agentic coding tasks.
Meta open-sourced Astryx, a React-based design system containing over 150 accessible components, built to be operated by both humans and AI agents.
Poolside released Laguna S 2.1, a 118B-parameter open-weight Mixture-of-Experts model optimized for agentic coding, featuring a 1M-token context window.
NVIDIA has released Cosmos 3 Edge, a 4-billion-parameter open world model designed for on-device robot and vision AI agent reasoning and action generation.
Google introduced three new Gemini models in the Flash tier—Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber—optimized for token efficiency and agentic workloads.
A community developer created a 657MB local thinking model by fine-tuning OpenBMB’s MiniCPM5-1B on Claude Fable 5 traces.
Alibaba's Tongyi Lab launched Qwen-Audio-3.0-TTS, a hosted text-to-speech model featuring Flash and Plus tiers and supporting 16 languages.
10 Open-Source No-Code AI Platforms for Building LLM Apps, RAG Systems, and AI Agents MarkTechPost
A comparison of Kimi K3, DeepSeek V4 Pro, and GLM-5.2, highlighting that while Kimi K3 leads in benchmarks, it is currently API-only, whereas the others offer open weights.
Feyn AI released SQRL, a family of text-to-SQL models that inspect databases before generating queries to resolve ambiguity.
Alibaba previewed Qwen3.8-Max, a 2.4 trillion-parameter multimodal model, at the World AI Conference, claiming performance second only to Fable 5.
Perplexity AI has released WANDR, an open benchmark designed to evaluate research agents on their ability to perform wide and deep information collection tasks.
Alibaba previewed Qwen3.8-Max, a 2.4 trillion-parameter multimodal model, during the World AI Conference in Shanghai.
Google Cloud's Always-On Memory Agent Replaces RAG and Embeddings With Continuous LLM Consolidation on Gemini 3.1 Flash-Lite MarkTechPost
Google Cloud's Always-On Memory Agent Replaces RAG and Embeddings With Continuous LLM Consolidation on Gemini 3.1 Flash-Lite MarkTechPost
NVIDIA released DeepStream 9.1, featuring 13 new agentic skills and enhanced multi-view 3D tracking capabilities for vision AI pipelines.
Sakana AI published research on Error Diffusion, a method for training Dale-compliant networks without backpropagation.
Google Cloud released an Always-On Memory Agent that uses a continuous process and SQLite storage instead of traditional RAG and embeddings, built on Gemini 3.1 Flash-Lite.
Zyphra released ZUNA1.1, an Apache 2.0 licensed EEG foundation model that now supports variable-length inputs ranging from 0.5 to 30 seconds.
Moonshot AI released Kimi K3, a 2.8-trillion-parameter open Mixture-of-Experts (MoE) model featuring a 1-million-token context window and Kimi Delta Attention.
NVIDIA released Nemotron 3 Embed, an open embedding collection that ranks #1 on the Retrieval Embedding Benchmark (RTEB).
Sakana AI published research on Error Diffusion, a method for training Dale-compliant networks without backpropagation that matches or exceeds traditional approaches in specific tasks.
NVIDIA released Nemotron 3 Embed, an open embedding collection featuring an 8B checkpoint that currently ranks first on the RTEB benchmark.
OpenAI Details GPT-Red: An Internal Automated Red-Teaming Model That Beat Human Red-Teamers 84% To 13% On Prompt Injection MarkTechPost
OpenAI detailed GPT-Red, an internal automated red-teaming model designed to identify prompt injection vulnerabilities by attacking OpenAI's own models.
Thinking Machines Lab released Inkling, a 975B-parameter open-weights multimodal MoE model featuring 41B active parameters and controllable thinking effort.
Moonshot AI released Kimi K3, a 2.8 trillion parameter open mixture-of-experts (MoE) model featuring native vision, a 1-million-token context window, and Kimi Delta Attention.
OpenAI detailed GPT-Red, an internal automated red-teaming model designed to find prompt injection vulnerabilities.
Google has released LiteRT.js, a JavaScript binding for LiteRT that allows for running .tflite models directly in web browsers using WebGPU.
The Soofi Consortium has released Soofi S 30B-A3B, an open hybrid Mamba-Transformer Mixture-of-Experts foundation model for German and English.
SpaceXAI open-sourced Grok Build, the Rust-based agent harness and tool layer supporting their coding CLI.
A comparison of Anthropic's Claude Sonnet 5, Sonnet 4.6, and Opus 4.8, focusing on agentic coding benchmarks, API pricing, and cost-performance trade-offs.
Mistral AI has launched Robostral Navigate, an 8B model designed to enable robots to navigate complex environments using only a single RGB camera.