AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,340 results
Model Releases

GEMSS: A Variational Method for Discovering Multiple Sparse Solutions in Classification and Regression Problems

DGX agent

arXiv:2602.08913v3 Announce Type: replace Abstract: In underdetermined regression and classification problems, multiple feature subsets often yield equivalent predictive performance. In applied settin

model-releasesarxiv-cs-lg
3 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Geographically Weighted Surrogate Models for Rapid Small-Area Chronic Disease Estimation

DGX agent

arXiv:2607.28655v1 Announce Type: cross Abstract: Small-area estimation (SAE) enables researchers and policymakers to identify spatial disparities in health outcomes, but survey-based SAE products car

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

GO-PRE: Goal-Oriented Next-Best-View Selection via Predictive Rendering Entropy for Active 3D Reconstruction

DGX agent

arXiv:2607.29037v1 Announce Type: new Abstract: Active 3D reconstruction relies on active view selection to maximize reconstruction fidelity under limited capture budgets. However, most existing metho

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Harnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration

DGX agent

arXiv:2607.29087v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in enterprise settings, yet individual models remain bounded by model-specific capability limitat

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

How to run big models on old hardware 30B at 22 tok/s on 6GB GPU and 16GB RAM

DGX agent

I have been working on this tool for months and there are a lot of new functionalities and tests that are going to be released in the next few weeks! The goal of the tool is to allow community members

model-releasesr-ollama
3 Aug 2026
Model Releases

Hy-MultiTurn: A Six-Dimensional Benchmark for Deep Multi-Turn Dialogue Understanding

DGX agent

arXiv:2607.29196v1 Announce Type: new Abstract: Long-running multi-turn interactions with chatbots and agents are now common, and a correct response often depends on remembering earlier details, track

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

I compared MinerU, Granite-Docling, and PaddleOCR-VL on 12 PDF-parsing capabilities using 6 document types

DGX agent

I tested them by sending the 6 documents, each meant to represent a different document type, through my own webapp and comparing every output against the source. All ran on the same L4 GPU. The docume

model-releasesr-localllama
3 Aug 2026
Model Releases

I gave five different local LLMs a town. They invented Facebook and a duck-based credit bureau. (MIT, self-hosted, you don't play it — you watch it)

DGX agent

Each villager in Pepperton is a different model — a mistral, a qwen3, a qwen2.5, a phi4-mini, a llama3.2 — because model families have genuinely different temperaments, and the friction between them i

model-releasesr-ollama
3 Aug 2026
Model Releases

Identifying Informative Environments for Cognition Parameter Inference via Bayesian Experimental Design

DGX agent

arXiv:2607.28894v1 Announce Type: new Abstract: Computational cognitive modeling seeks to infer latent cognitive mechanisms underlying observed behavior. Bayesian inverse planning provides a principle

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Inference-time Trajectory Optimization for Structure-Preserving Manga Image Editing

DGX agent

arXiv:2603.27790v2 Announce Type: replace Abstract: We present a lightweight, training-free trajectory correction method that adapts a pretrained image editing model to each input manga image using on

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

InferQ: A Database-Oriented Benchmark for Quantum Circuits Simulation

DGX agent

arXiv:2607.29134v1 Announce Type: cross Abstract: Recent work suggests that relational database management systems (RDBMSs) can execute quantum circuit simulation by compiling the simulation into SQL

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Is It Time for the Renaissance of Salient Object Detection in the Era of MLLMs?

DGX agent

arXiv:2607.29222v1 Announce Type: new Abstract: The zero-shot capabilities of multimodal large language models (MLLMs) are pushing salient object detection (SOD) beyond task-specific supervision. To d

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

KAT Coder 2.5 dev: Do yourself a favor and try it!

DGX agent

It is so good! I don't know why there aren't more people talking about it. Fewer tokens, faster and more accurate than Qwen 3.6 35b a3b. On my setup it's nearly as good as 27b, but 5x faster. And it c

model-releasesr-localllama
3 Aug 2026
Model Releases

Language Models Agree With Each Other, Not With Readers

DGX agent

arXiv:2607.29274v1 Announce Type: cross Abstract: Claims that language models homogenise are usually measured against human judgements collected for the study, which makes the human side an artifact o

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

LARA: Lightweight Adapters in the Residual Stream for Composable Adaptation and Alignment

DGX agent

arXiv:2607.28669v1 Announce Type: new Abstract: We present LARA (Lightweight Additive Residual Adaptation), a method for efficient adaptation that operates in the residual stream of a frozen model rat

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Latent Sculpting for Zero-Shot Generalization: A Manifold Learning Approach to Out-of-Distribution Anomaly Detection

DGX agent

arXiv:2512.22179v3 Announce Type: replace Abstract: Detecting previously unseen attacks remains a major challenge for machine learning-based intrusion detection systems. Deep models trained on network

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

LayoutBench: Performance Benchmarking of Cloud Storage Layouts for Multimedia Data

DGX agent

arXiv:2607.28880v1 Announce Type: cross Abstract: Modern multimedia machine learning workloads increasingly store large-scale datasets in cloud object storage services such as AWS S3. How these sample

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

Learning Optimal Dynamic Matching via Graph Neural Networks

DGX agent

arXiv:2607.28925v1 Announce Type: new Abstract: Dynamic matching markets require decisions about whom to match and when: matching now yields value but removes participants who may create better future

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback

DGX agent

arXiv:2607.29559v1 Announce Type: new Abstract: Reinforcement Learning (RL) systems are typically trained using a single, well-specified scalar reward function. However, real-world decision-making tas

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

🔥Let's talk about Qwen! #AMA

DGX agent

The post announces an “Ask Me Anything” (AMA) about Qwen, the Alibaba‑developed foundation model. It introduces the Qwen Foundation Model Team and directs readers to their GitHub repository @QwenDevs,

model-releasesqwen--x
3 Aug 2026
Model Releases

Linear Proposal Operators and Stochastic Search Geometry in SOMA and Differential Evolution

DGX agent

arXiv:2607.29228v1 Announce Type: cross Abstract: Swarm and evolutionary algorithms are usually analyzed as complete procedural systems in which nonlinear selection, replacement, and adaptation obscur

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Ling-3.0-flash is another potential model to test before qwen3.8 27b

DGX agent

I tested Ling-3.0-flash with hard bugs and it fixed bugs that qwen3.6-27b could not. This models speed faster than deepseek v4 flash but almost the same level as (old) deepseek v4 flash. Note: hard bu

model-releasesr-localllama
3 Aug 2026
Model Releases

Live in Command Code!🎉

DGX agent

Live in Command Code!🎉 Qwen 3.8 Max is now live in Command Code Go. … and it's going open weight!! most importantly, next week open-weights of Qwen3.8-Max and Qwen3.8-27B will be released! 🔹 2.4T para

model-releasesqwen--x
3 Aug 2026
Model Releases

Locally Consistent Transductive Information Maximization for Few-Shot Remote Sensing Scene Classification

DGX agent

arXiv:2607.29192v1 Announce Type: new Abstract: Remote sensing scene classification is increasingly relying on foundation models pre-trained on large-scale Earth-observation data. Moreover, transducti

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Looks Right, Works Right: A Project-Level Benchmark for Multi-Screen Mobile App Generation

DGX agent

arXiv:2607.28645v1 Announce Type: cross Abstract: Recent multimodal large language models can convert visual designs directly into executable code, but real mobile products require multiple screenshot

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face Hack

DGX agent

LWiAI Podcast #253 (July 29 2026) covers a roundup of recent AI developments: Anthropic introduced Claude Opus 5 with Fable‑like capabilities; Google released Gemini 3.6/3.5 “Flash” variants and a cyb

model-releaseslast-week-in-ai
3 Aug 2026
Model Releases

M3-DuplexBench: A Multi-Turn, Multilingual, Multidomain Benchmark for Full-Duplex Spoken Dialogue Models

DGX agent

arXiv:2607.29125v1 Announce Type: new Abstract: Full-duplex spoken dialogue systems (FDSDSs) can listen while speaking, enabling natural behaviors such as smooth turn-taking, backchannel handling, and

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

DGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Matterhorn: Masked Time-to-First-Spike Encoding by Reassigning the Silent State for Sparse and Energy-Efficient Spiking Transformers

DGX agent

arXiv:2601.22876v2 Announce Type: replace Abstract: Spiking neural networks (SNNs) promise energy-efficient inference for large language models (LLMs), yet most reported savings rely on compute-operat

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

MDIR: A Task-Manifold Impedance Retargeting Method for Contact-Rich Teleoperation

DGX agent

arXiv:2607.29271v1 Announce Type: new Abstract: Fixed Cartesian impedance makes contact-rich teleoperation demonstrations practical, but gains that secure progress and contact support also determine i

model-releasesarxiv-cs-ro
3 Aug 2026
Model Releases

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also go…

DGX agent

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉 Qwen3.8-Max, a new bar for

model-releasesqwen--x
3 Aug 2026
Model Releases

MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft

DGX agent

arXiv:2607.29218v1 Announce Type: new Abstract: With the prosperity of the large language models (LLMs), it has become an interesting topic: how do LLM-based agents work in Minecraft? Unfortunately, m

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents

DGX agent

arXiv:2607.29002v1 Announce Type: new Abstract: Online shoppers increasingly turn to AI shopping assistants, using images and multi-turn dialogue to express and refine product needs that are difficult

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

DGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Models

DGX agent

arXiv:2607.29431v1 Announce Type: new Abstract: Large language models increasingly generate optimization models from natural language, but existing evaluation often reduces a generated model and its g

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

MoPET: Parameter-Efficient Mixture-of-Experts for Unified Medical Image Classification

DGX agent

arXiv:2607.29462v1 Announce Type: cross Abstract: Adapting deep learning models to profound clinical heterogeneity typically relies on parameter-efficient fine-tuning (PEFT) to avoid the severe overfi

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

MoRoute: Dynamic Routing for In-Context Multimodal Video Generation

DGX agent

arXiv:2607.29545v1 Announce Type: new Abstract: Multimodal video generation aims to generate and edit videos conditioned on arbitrary combinations of text, images, and videos within a single model, al

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO

DGX agent

arXiv:2607.28679v1 Announce Type: new Abstract: Multi-agent planning problems arise in a variety of engineering applications, such as multi-robot wildfire fighting and unmanned aerial inspection in fa

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Nate is right (full context can lead to a degradation of chats in several different ways) but you can't ask the AI about stuff like this as …

DGX agent

Nate is right (full context can lead to a degradation of chats in several different ways) but you can't ask the AI about stuff like this as they have bad self-knowledge A good approach is to compact t

model-releasesethan-mollick--x
3 Aug 2026
Model Releases

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up …

DGX agent

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up with 27.3% of the money a human makes. MerchantBench is a 36

model-releasesdair-ai--x
3 Aug 2026
Model Releases

'Not in My Backyard': LLMs Uncover Online and Offline Social Biases Against Homelessness

DGX agent

arXiv:2508.13187v5 Announce Type: replace-cross Abstract: Homelessness is a persistent social challenge, impacting millions worldwide. Over 876,000 people experiencing homelessness (PEH) were recorded

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

NousResearch keeps doing things on hermes

DGX agent

Has anyone followed nousresearch work on Hermes? I mean we are Q3 2026. We have some crazy models trickling down from HGX territory to multi gpu workstation. And we have nousresearch deploying the 0.2

model-releasesr-localllama
3 Aug 2026
Model Releases

On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness

DGX agent

arXiv:2607.29062v1 Announce Type: new Abstract: Model capabilities have improved in large part due to scaling chain of thought. This has been a promising development for AI safety--where models verbal

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair

DGX agent

arXiv:2607.28877v1 Announce Type: cross Abstract: Verification consumes the majority of modern chip design effort, yet the formal verification tools that provide mathematical guarantees of correctness

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems

DGX agent

arXiv:2607.28629v1 Announce Type: new Abstract: The rapid transition from reactive large language models (LLMs) to persistent, action-capable systems has exposed critical gaps in the architectural und

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

OSAGEN: Object-Aware Mask Priors and Multistage Decoupled Diffusion for Industrial Anomaly Generation

DGX agent

arXiv:2607.29533v1 Announce Type: new Abstract: Industrial anomaly detection and localization are limited by scarce real anomalies and pixel-level annotations, a bottleneck that synthetic image-mask p

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

OSEF: One-Step Evidence Fusion for Cross-Video Scene Procedure Planning

DGX agent

arXiv:2607.29401v1 Announce Type: new Abstract: Video Scene Procedure Planning (VSPP) supplies the target start-goal observations in advance, leaving open how a planner should act when the evidence mu

model-releasesarxiv-cs-cv
3 Aug 2026
Model Releases

PARALLEL: A Prefrontal-Aligned Reinforcement inspired Approach for Language-Model Learning under Explicit Limits

DGX agent

arXiv:2607.28982v1 Announce Type: cross Abstract: Recent language models achieve strong performance across a variety of tasks, but conventional adaptation applies updates uniformly across training sam

model-releasesarxiv-cs-ai
3 Aug 2026
← Previous
1…5253545556…466
Next →