-
JustFit: 200K-Token LLM Serving on a 24 GiB Laptop with Just-in-Time State Management
ArXiv · AI/CL/LG
·
JustFit enables 200K-token local LLM serving on a 24 GiB MacBook through just-in-time execution-state management.
-
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search
HF Daily Papers
·
ZGCM-1 is a fully open 7B foundation model trained from scratch for math and agentic search with 256K context.
-
Don't sleep on wrapture
Simon Willison
·
Wrapture is a new Python monkey-patching library aimed at testing and observability use cases.
-
Building py-kvcache: A Performance Characterization of External KV Caching for vLLM with NVMe SSDs
ArXiv · AI/CL/LG
·
py-kvcache adds scheduler-aware NVMe KV-cache offload for vLLM, improving long-context serving performance under realistic workloads.
-
Shannon 3.0 AI pentester gets through Aikido and XBOW
TestingCatalog
·
Keygraph released an open-source AI pentester that analyzes source architecture before exploiting live applications.
-
IFM releases K2 Horizon: 6 models with full training record
TestingCatalog
·
IFM released six open K2 Horizon checkpoints up to 375B with code, training logs, and evaluations for reproducible model development.
-
OpenWAM: An Open, Modular Exploration Towards Systematic World-Action Model Pretraining
HF Daily Papers
·
OpenWAM provides an open modular stack for studying world-action model pretraining, evaluation, inference, and deployment.
-
VidaForge: Open Research Infrastructure for Video Pretraining Data Recipes
HF Daily Papers
·
VIDAFORGE opens executable video pretraining data recipes so researchers can compare pipeline choices and dataset provenance.
-
K2 Horizon: A connected fleet of six open models
HN · Frontpage AI
·
K2 Horizon released six connected open models, drawing strong builder interest around a coordinated open-weights model family.
-
WebLLM: high-performance in-browser LLM inference engine
HN · LLMs
·
WebLLM brings high-performance LLM inference directly into browsers, making local client-side AI applications more practical.
-
GRADSOLVE: fast exact gradients for ODE ensembles on GPUs
ArXiv · AI/CL/LG
·
GRADSOLVE is an open-source JAX library for fast reverse-mode gradients over GPU ODE ensembles.
-
shadcn-ui/lint
GitHub · Agents repos
·
shadcn-ui released an agent-oriented linter for Tailwind design systems, letting teams encode UI rules that coding agents can verify.
-
Relational-Core Graph Analytics Querying graphs at SQL scale, and why the node/edge model is a performance tax, not a truer picture of connected data
ArXiv · AI/CL/LG
·
ClickGraph and DeltaGraph translate Cypher onto relational engines, challenging native graph databases for analytical workloads.
-
OpenClaw releases OpenClaw 2.0, its largest update to date built by 933 contributors, with a simplified installation process, a rebuilt browser app, and more (Hannes Rudolph/OpenClaw Blog)
Techmeme
·
OpenClaw 2.0 ships a major community-built update with easier setup and a rebuilt browser app for the open-source agent tool.
-
tigerless-labs/agent-memory
GitHub · LLM repos
·
Agent-memory offers a local-first long-term memory runtime for AI agents using Markdown, ranked retrieval, and sleep-time management.
-
Introducing Hy4 Preview
Simon Willison
·
Tencent released Hy4 Preview, a 770B-parameter open-weight text LLM with 49B active parameters and a 1M-token context window.
-
E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation
HF Daily Papers
·
E-Commerce Bench is an open-source benchmark for year-long autonomous business operations with negotiation and dynamic events.
-
Dr. Claw: An AI Scientist Workspace for Vibe Research
HF Daily Papers
·
Dr. Claw is an open-source AI scientist workspace for auditable, human-in-the-loop research workflows across coding-agent executors.
-
Tencent released open-source Hy4 preview model
TestingCatalog
·
Tencent released an open-source Hy4 preview with 770B parameters and a 1M-token context window for coding, office, games, and research.
-
2akouwu/reverify
GitHub · LLM repos
·
Reverify is an MCP server and CLI that verifies LLM-generated reverse-engineering claims against binary bytes for more reliable analysis.
-
GLM-5.3 is now open-weight
HN · Hugging Face Models
·
GLM-5.3 was released as open weights, giving builders access to a notable new Chinese frontier-style model.
-
Human-Agent-Society/reef
GitHub · LLM repos
·
Reef offers continual-learning infrastructure for self-improving agents, a narrow but relevant developer tool for agent training workflows.
-
Anthropic releases Model Hardware Standard, a framework to help AI agents use physical systems like microscopes, quantum computing hardware, and robot arms (Will Knight/Wired)
Techmeme
·
Anthropic released a hardware-control standard aimed at letting AI agents safely operate scientific instruments and robotics systems.
-
Alibaba releases Qwen3.8-Flash, an open-weight, 125B-parameter model built on its next-gen Qwen 4 architecture, saying it rivals Opus 4.6 and V4-Flash (Luz Ding/Bloomberg)
Techmeme
·
Alibaba released an open-weight 125B Qwen model on its next-generation architecture with claimed frontier-adjacent performance.
-
Z.ai launches GLM-5.3-Flash under MIT license
TestingCatalog
·
Z.ai released open MIT-licensed GLM-5.3-Flash weights for multimodal reasoning, coding, and local deployment.
-
Z.ai confirms Ox Alpha is a new iteration of its GLM series and says it will release the weights for it tonight; Ox Alpha topped OpenRouter's leaderboard (Luz Ding/Bloomberg)
Techmeme
·
Z.ai confirmed Ox Alpha as a new GLM model and plans to release weights after it topped OpenRouter's leaderboard.
-
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090
HF Daily Papers
·
Puro-2B releases small open models and a low-cost pretraining recipe using consumer RTX 5090 GPUs.
-
Player-YN/PawWork_ZhuaZhua
GitHub · LLM repos
·
PawWork is an open-source Chrome agent that turns selected web content into editable office files without a server.
-
Nanako0129/sepia
GitHub · LLM repos
·
An open-source writing skill adapts AI-assisted prose toward fiction and professional style rules, with modest early GitHub traction.
-
Hugging Face is selling a cute $399 open-source duck robot, Microduck
TechCrunch AI
·
Hugging Face launched a $399 open-source home robot for developers to train and experiment with embodied AI.
-
LAION-BVD: A 10-Million-Hour Open Video Dataset for Multimodal Pre-training
HF Daily Papers
·
LAION-BVD releases a 10-million-hour open video dataset aimed at large-scale multimodal pretraining across video, audio, and image tasks.
-
TorchMorph: CUDA-accelerated Morphological Transforms
HF Daily Papers
·
TorchMorph adds CUDA-accelerated PyTorch morphology and distance-transform operators for GPU-native vision and mask-processing training loops.
-
CyberFactory: Scaling Cyber Security Capabilities with Instances from the Wild
HF Daily Papers
·
CyberFactory provides an open framework for building cybersecurity training data, agent trajectories, and models across offensive and defensive tasks.
-
Prime Agent: A Self-Improving RLM Harness
HF Daily Papers
·
Prime Agent is an open-source long-horizon agent harness with persistent execution, memory, subagents, and human inspection.
-
rome-os/rome
GitHub · LLM repos
·
Rome is an open-source TypeScript agent OS for recursive agent workflows, crossing the launch notability threshold with early GitHub traction.
-
Ornith-1.5 open models launch in 397B, 35B, and 9 B sizes.
TestingCatalog
·
Ornith-1.5 ships open models up to 397B with self-improving task generation and strong coding and reasoning benchmark claims.
-
Mojo🔥 is now open source
Simon Willison
·
Modular open-sourced Mojo's compiler and toolchain under Apache 2, making its AI-focused systems language independently inspectable and extensible.
-
Block releases Berd, a desktop app it built to give its employees a single environment for working with AI agents across different models, under Apache 2.0 (Carl Franzen/VentureBeat)
Techmeme
·
Block open-sourced Berd, a desktop workspace for running AI agents across multiple models with local conversation history.
-
Institutional Books - Enriched Text: A customizable multilingual open-source pipeline for denoising, deduplicating, and annotating OCR text at scale
ArXiv · AI/CL/LG
·
Institutional Books Enriched Text provides an open-source multilingual pipeline for denoising, deduplicating, and annotating large OCR book corpora.