AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,323 results
3 Aug 2026

FairFund-Bench: Evaluating Distributive Bias in LLM Resource Allocation

Model ReleasesDGX agent

arXiv:2607.28934v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly involved in the distribution of scarce resources, raising concerns about biased allocations based on cha

Fast Rates for Swap-Agnostic Learning of Proper Losses

Model ReleasesDGX agent

arXiv:2607.28856v1 Announce Type: new Abstract: Swap-agnostic learning strengthens classical agnostic learning by allowing the comparator to select a different hypothesis on each level set of the lear

Federated Foundation Models Fine-Tuning with Heterogeneous Compressed Clients

Model ReleasesDGX agent

arXiv:2607.29071v1 Announce Type: cross Abstract: Federated learning of foundation models faces a fundamental resource-asymmetry challenge: the institutions holding the most valuable domain-specific d


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FlexComposer: Unified Video Compositing from Images to Dynamic Footage with Flexible Trajectory Control

Model ReleasesDGX agent

arXiv:2607.29627v1 Announce Type: new Abstract: Generative video compositing, which involves inserting external assets seamlessly into existing video sequences, is essential for content creation and v

FriendBench: Benchmarking Dyadic Familiarity Inference in Humans and Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2607.29602v1 Announce Type: cross Abstract: Reading a social situation often depends on behavior, not words alone. We introduce FriendBench, a benchmark for inferring whether two people are alre

Frontier AI performance across the business disciplines: a case-grounded benchmark of knowledge work and analytical reasoning

Model ReleasesDGX agent

arXiv:2607.16057v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are improving rapidly as reflected in benchmark scores, yet these AI benchmarks largely test capabilities such as

Frugal Bayesian Optimization: Scalable Surrogates for Data- and Resource-Limited Discovery

Model ReleasesDGX agent

arXiv:2607.29225v1 Announce Type: new Abstract: Bayesian Optimization (BO) is widely adopted for data-efficient optimization in scientific and engineering applications, yet its computational cost is r

GEMSS: A Variational Method for Discovering Multiple Sparse Solutions in Classification and Regression Problems

Model ReleasesDGX agent

arXiv:2602.08913v3 Announce Type: replace Abstract: In underdetermined regression and classification problems, multiple feature subsets often yield equivalent predictive performance. In applied settin

Geographically Weighted Surrogate Models for Rapid Small-Area Chronic Disease Estimation

Model ReleasesDGX agent

arXiv:2607.28655v1 Announce Type: cross Abstract: Small-area estimation (SAE) enables researchers and policymakers to identify spatial disparities in health outcomes, but survey-based SAE products car

GO-PRE: Goal-Oriented Next-Best-View Selection via Predictive Rendering Entropy for Active 3D Reconstruction

Model ReleasesDGX agent

arXiv:2607.29037v1 Announce Type: new Abstract: Active 3D reconstruction relies on active view selection to maximize reconstruction fidelity under limited capture budgets. However, most existing metho

Harnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration

Model ReleasesDGX agent

arXiv:2607.29087v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in enterprise settings, yet individual models remain bounded by model-specific capability limitat

How to run big models on old hardware 30B at 22 tok/s on 6GB GPU and 16GB RAM

Model ReleasesDGX agent

I have been working on this tool for months and there are a lot of new functionalities and tests that are going to be released in the next few weeks! The goal of the tool is to allow community members

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

Model ReleasesDGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

I compared MinerU, Granite-Docling, and PaddleOCR-VL on 12 PDF-parsing capabilities using 6 document types

Model ReleasesDGX agent

I tested them by sending the 6 documents, each meant to represent a different document type, through my own webapp and comparing every output against the source. All ran on the same L4 GPU. The docume

I gave five different local LLMs a town. They invented Facebook and a duck-based credit bureau. (MIT, self-hosted, you don't play it — you watch it)

Model ReleasesDGX agent

Each villager in Pepperton is a different model — a mistral, a qwen3, a qwen2.5, a phi4-mini, a llama3.2 — because model families have genuinely different temperaments, and the friction between them i

Identifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design

Model ReleasesDGX agent

arXiv:2607.28894v1 Announce Type: new Abstract: Computational cognitive modeling seeks to infer latent cognitive mechanisms underlying observed behavior. Bayesian inverse planning provides a principle

Inference-time Trajectory Optimization for Structure-Preserving Manga Image Editing

Model ReleasesDGX agent

arXiv:2603.27790v2 Announce Type: replace Abstract: We present a lightweight, training-free trajectory correction method that adapts a pretrained image editing model to each input manga image using on

InferQ: A Database-Oriented Benchmark for Quantum Circuits Simulation

Model ReleasesDGX agent

arXiv:2607.29134v1 Announce Type: cross Abstract: Recent work suggests that relational database management systems (RDBMSs) can execute quantum circuit simulation by compiling the simulation into SQL

Is It Time for the Renaissance of Salient Object Detection in the Era of MLLMs?

Model ReleasesDGX agent

arXiv:2607.29222v1 Announce Type: new Abstract: The zero-shot capabilities of multimodal large language models (MLLMs) are pushing salient object detection (SOD) beyond task-specific supervision. To d

KAT Coder 2.5 dev: Do yourself a favor and try it!

Model ReleasesDGX agent

It is so good! I don't know why there aren't more people talking about it. Fewer tokens, faster and more accurate than Qwen 3.6 35b a3b. On my setup it's nearly as good as 27b, but 5x faster. And it c

Language Models Agree With Each Other, Not With Readers

Model ReleasesDGX agent

arXiv:2607.29274v1 Announce Type: cross Abstract: Claims that language models homogenise are usually measured against human judgements collected for the study, which makes the human side an artifact o

LARA: Lightweight Adapters in the Residual Stream for Composable Adaptation and Alignment

Model ReleasesDGX agent

arXiv:2607.28669v1 Announce Type: new Abstract: We present LARA (Lightweight Additive Residual Adaptation), a method for efficient adaptation that operates in the residual stream of a frozen model rat

Latent Sculpting for Zero-Shot Generalization: A Manifold Learning Approach to Out-of-Distribution Anomaly Detection

Model ReleasesDGX agent

arXiv:2512.22179v3 Announce Type: replace Abstract: Detecting previously unseen attacks remains a major challenge for machine learning-based intrusion detection systems. Deep models trained on network

LayoutBench: Performance Benchmarking of Cloud Storage Layouts for Multimedia Data

Model ReleasesDGX agent

arXiv:2607.28880v1 Announce Type: cross Abstract: Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3. How these sample

Learning Optimal Dynamic Matching via Graph Neural Networks

Model ReleasesDGX agent

arXiv:2607.28925v1 Announce Type: new Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback

Model ReleasesDGX agent

arXiv:2607.29559v1 Announce Type: new Abstract: Reinforcement Learning (RL) systems are typically trained using a single, well-specified scalar reward function. However, real-world decision-making tas

🔥Let's talk about Qwen! #AMA

Model ReleasesDGX agent

The post announces an “Ask Me Anything” (AMA) about Qwen, the Alibaba‑developed foundation model. It introduces the Qwen Foundation Model Team and directs readers to their GitHub repository @QwenDevs,

Linear Proposal Operators and Stochastic Search Geometry in SOMA and Differential Evolution

Model ReleasesDGX agent

arXiv:2607.29228v1 Announce Type: cross Abstract: Swarm and evolutionary algorithms are usually analyzed as complete procedural systems in which nonlinear selection, replacement, and adaptation obscur

Ling-3.0-flash is another potential model to test before qwen3.8 27b

Model ReleasesDGX agent

I tested Ling-3.0-flash with hard bugs and it fixed bugs that qwen3.6-27b could not. This models speed faster than deepseek v4 flash but almost the same level as (old) deepseek v4 flash. Note: hard bu

Live in Command Code!🎉

Model ReleasesDGX agent

Live in Command Code!🎉 Qwen 3.8 Max is now live in Command Code Go. … and it's going open weight!! most importantly, next week open-weights of Qwen3.8-Max and Qwen3.8-27B will be released! 🔹 2.4T para

Locally Consistent Transductive Information Maximization for Few-Shot Remote Sensing Scene Classification

Model ReleasesDGX agent

arXiv:2607.29192v1 Announce Type: new Abstract: Remote sensing scene classification is increasingly relying on foundation models pre-trained on large-scale Earth-observation data. Moreover, transducti

Looks Right, Works Right: A Project-Level Benchmark for Multi-Screen Mobile App Generation

Model ReleasesDGX agent

arXiv:2607.28645v1 Announce Type: cross Abstract: Recent multimodal large language models can convert visual designs directly into executable code, but real mobile products require multiple screenshot

LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face Hack

Model ReleasesDGX agent

LWiAI Podcast #253 (July 29 2026) covers a roundup of recent AI developments: Anthropic introduced Claude Opus 5 with Fable‑like capabilities; Google released Gemini 3.6/3.5 “Flash” variants and a cyb

M3-DuplexBench: A Multi-Turn, Multilingual, Multidomain Benchmark for Full-Duplex Spoken Dialogue Models

Model ReleasesDGX agent

arXiv:2607.29125v1 Announce Type: new Abstract: Full-duplex spoken dialogue systems (FDSDSs) can listen while speaking, enabling natural behaviors such as smooth turn-taking, backchannel handling, and

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

Model ReleasesDGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

Matterhorn: Masked Time-to-First-Spike Encoding by Reassigning the Silent State for Sparse and Energy-Efficient Spiking Transformers

Model ReleasesDGX agent

arXiv:2601.22876v2 Announce Type: replace Abstract: Spiking neural networks (SNNs) promise energy-efficient inference for large language models (LLMs), yet most reported savings rely on compute-operat

MDIR: A Task-Manifold Impedance Retargeting Method for Contact-Rich Teleoperation

Model ReleasesDGX agent

arXiv:2607.29271v1 Announce Type: new Abstract: Fixed Cartesian impedance makes contact-rich teleoperation demonstrations practical, but gains that secure progress and contact support also determine i

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also go…

Model ReleasesDGX agent

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉 Qwen3.8-Max, a new bar for

MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft

Model ReleasesDGX agent

arXiv:2607.29218v1 Announce Type: new Abstract: With the prosperity of the large language models (LLMs), it has become an interesting topic: how do LLM-based agents work in Minecraft? Unfortunately, m

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

Model ReleasesDGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

Model ReleasesDGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models

Model ReleasesDGX agent

arXiv:2607.29431v1 Announce Type: new Abstract: Large language models increasingly generate optimization models from natural language, but existing evaluation often reduces a generated model and its g

MoPET: Parameter-Efficient Mixture-of-Experts for Unified Medical Image Classification

Model ReleasesDGX agent

arXiv:2607.29462v1 Announce Type: cross Abstract: Adapting deep learning models to profound clinical heterogeneity typically relies on parameter-efficient fine-tuning (PEFT) to avoid the severe overfi

MoRoute: Dynamic Routing for In-Context Multimodal Video Generation

Model ReleasesDGX agent

arXiv:2607.29545v1 Announce Type: new Abstract: Multimodal video generation aims to generate and edit videos conditioned on arbitrary combinations of text, images, and videos within a single model, al

Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO

Model ReleasesDGX agent

arXiv:2607.28679v1 Announce Type: new Abstract: Multi-agent planning problems arise in a variety of engineering applications, such as multi-robot wildfire fighting and unmanned aerial inspection in fa

Nate is right (full context can lead to a degradation of chats in several different ways) but you can't ask the AI about stuff like this as …

Model ReleasesDGX agent

Nate is right (full context can lead to a degradation of chats in several different ways) but you can't ask the AI about stuff like this as they have bad self-knowledge A good approach is to compact t

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up …

Model ReleasesDGX agent

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up with 27.3% of the money a human makes. MerchantBench is a 36

'Not in My Backyard': LLMs Uncover Online and Offline Social Biases Against Homelessness

Model ReleasesDGX agent

arXiv:2508.13187v5 Announce Type: replace-cross Abstract: Homelessness is a persistent social challenge, impacting millions worldwide. Over 876,000 people experiencing homelessness (PEH) were recorded

NousResearch keeps doing things on hermes

Model ReleasesDGX agent

Has anyone followed nousresearch work on Hermes? I mean we are Q3 2026. We have some crazy models trickling down from HGX territory to multi gpu workstation. And we have nousresearch deploying the 0.2

On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness

Model ReleasesDGX agent

arXiv:2607.29062v1 Announce Type: new Abstract: Model capabilities have improved in large part due to scaling chain of thought. This has been a promising development for AI safety--where models verbal

Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair

Model ReleasesDGX agent

arXiv:2607.28877v1 Announce Type: cross Abstract: Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness

OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems

Model ReleasesDGX agent

arXiv:2607.28629v1 Announce Type: new Abstract: The rapid transition from reactive large language models (LLMs) to persistent, action-capable systems has exposed critical gaps in the architectural und

OSAGEN: Object-Aware Mask Priors and Multistage Decoupled Diffusion for Industrial Anomaly Generation

Model ReleasesDGX agent

arXiv:2607.29533v1 Announce Type: new Abstract: Industrial anomaly detection and localization are limited by scarce real anomalies and pixel-level annotations, a bottleneck that synthetic image-mask p

OSEF: One-Step Evidence Fusion for Cross-Video Scene Procedure Planning

Model ReleasesDGX agent

arXiv:2607.29401v1 Announce Type: new Abstract: Video Scene Procedure Planning (VSPP) supplies the target start-goal observations in advance, leaving open how a planner should act when the evidence mu

PARALLEL: A Prefrontal-Aligned Reinforcement inspired Approach for Language-Model Learning under Explicit Limits

Model ReleasesDGX agent

arXiv:2607.28982v1 Announce Type: cross Abstract: Recent language models achieve strong performance across a variety of tasks, but conventional adaptation applies updates uniformly across training sam

Parameter-Efficient Fine-Tuning for Spiking Point Cloud Models

Model ReleasesDGX agent

arXiv:2607.29048v1 Announce Type: new Abstract: Spiking Neural Networks (SNNs) offer energy-efficient solutions for point cloud analysis on resource-constrained devices through event-driven computatio

Parameter-Free Heavy-Tailed Bandits

Model ReleasesDGX agent

arXiv:2607.29460v1 Announce Type: new Abstract: Heavy-tailed distributions arise naturally in sequential decision-making problems such as financial investment, online advertising, and network manageme

Paris: A Decentralized Trained Open-Weight Diffusion Model

Model ReleasesDGX agent

arXiv:2510.03434v3 Announce Type: replace-cross Abstract: We present Paris, the first publicly released diffusion model pre-trained entirely through decentralized computation. Paris demonstrates that

Patch-Based 3D Variational Autoencoder for Super-Resolution of Turbulent Channel Flow

Model ReleasesDGX agent

arXiv:2507.22082v2 Announce Type: replace-cross Abstract: Direct numerical simulation (DNS) accurately resolves all spatio-temporal scales of wall-bounded turbulence but becomes prohibitively expensiv

PluRel-to-RDB-PFN: Schema-Guided Synthetic Relational Pretraining

Model ReleasesDGX agent

arXiv:2607.29129v1 Announce Type: new Abstract: Relational Foundation Models (RFMs) require large-scale synthetic relational databases for pretraining, but existing approaches tightly couple data gene

← Previous
1…4142434445…373
Next →