AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,553 results
15 Jul 2026

Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters

Model ReleasesDGX agent

MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models

ResearchDGX agent

arXiv:2602.02244v3 Announce Type: replace-cross Abstract: The standard post-training recipe for large reasoning models, supervised fine-tuning followed by reinforcement learning (SFT-then-RL), may lim

Graph Feedback Controls Consensus and Clique Formation in Open-Weight Language-Model Populations

Local AiDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.12077v1 Announce Type: new Abstract: Multi-agent language-model systems increasingly route local interactions, yet the runtime interaction graph is often treated as an implementation detail

Hy-Embodied-VLM-1.0: Efficient Physical-World Agents

Model ReleasesDGX agent

arXiv:2607.12894v1 Announce Type: new Abstract: Building capable embodied agents requires not only multimodal perception and understanding, but also agentic capabilities for reasoning about actions, a

LP Mining with LP2Graph: A Use Case for Railway Rescheduling

ResearchDGX agent

arXiv:2607.11980v1 Announce Type: new Abstract: Like many optimization-driven domains, railway rescheduling relies on Mixed-Integer Linear Programming (MILP), yet the field's modeling knowledge is sca

PolarBM: Complex-valued Boltzmann Machine for Modeling Audio Signals in Polar and Log-polar Coordinates

ResearchDGX agent

arXiv:2607.12417v1 Announce Type: new Abstract: Although vast amounts of data, such as audio signal spectra, are naturally represented using complex numbers, conventional machine learning methods ofte

SlimPer: Make Personalization Model Slim and Smart

ResearchDGX agent

arXiv:2607.12281v1 Announce Type: cross Abstract: Transformer-style architectures are increasingly adopted for industrial recommendation systems, yet they inherit a design premise misaligned with the

SymbOmni: Evolving Agentic Omni Models via Symbolic Concept Learning

AgentsDGX agent

arXiv:2607.12042v1 Announce Type: new Abstract: Visual generation is increasingly ubiquitous in diverse domains, from text-to-image/video synthesis to multimodal interactive creation. Yet prevailing m

Watermark Forensics for Generative Models: An Information-Theoretic Perspective

SafetyDGX agent

arXiv:2607.13003v1 Announce Type: cross Abstract: A watermark in a generative model's output is usually asked only whether a text is machine-made. The same mark can do more: attribute it to the user w

You can use OpenCode Desktop with Ollama! Try it with the top open models!

Local AiDGX agent

You can use OpenCode Desktop with Ollama! Try it with the top open models! Introducing Tabs OpenCode Desktop is now built around tabs. Start a new session in a tab, or open an existing session from an

13 Jul 2026

Language models and coding agents are great, but there is more to life, and more to AI, than just LLM agents.

ResearchDGX agent

Hardmaru posted a tweet on July 13, 2026 stating that while language models and coding agents are valuable, “there is more to life, and more to AI, than just LLM agents.” The tweet received 61.1 k vie

Ollama's @jmorgan is live on Yahoo Finance to talk about open models. https://finance.yahoo.com/live/

Local AiDGX agent

Ollama’s partner @jmorgan hosted a live session on Yahoo Finance on July 13, 2026 at 8:30 PM, where they discussed open‑model initiatives. The event was streamed in real time and attracted 16.3 K view

12 Jul 2026

Fable gets another bump

Model ReleasesDGX agent

One of the consequences of GPT-5.6 Sol being clearly a Fable/Mythos class model is that Anthropic have, once again, bumped the date that Fable stops being available in their Claude Max plans: We're ex

10 Jul 2026

COALA: Robust Contextualized Speech-augmented Language Modeling for ASR via Contrastive Regularizer and Biasing Score Estimation

Model ReleasesDGX agent

arXiv:2607.08117v1 Announce Type: new Abstract: Contextual biasing seeks to integrate external knowledge into automatic speech recognition (ASR) systems to accurately recognize domain-specific entitie

CommuniWave:A Machine Learning Model for Quantifying the Degree of Temporary Informal Behavior in Urban Communities

ResearchDGX agent

arXiv:2607.08554v1 Announce Type: new Abstract: For urban managers and designers, improving the functional attributes of urban communities to enhance territorial resilience in the face of complexity a

Concretized Proposition Prompting Resolves Composition-Knowledge Dichotomy in Large Language Models

Model ReleasesDGX agent

arXiv:2607.08018v1 Announce Type: new Abstract: LLMs often struggle to balance compositionality with knowledgeability, a challenge we define as Composition-Knowledge Dichotomy. To address this, we pro

Dropping Just a Handful of Preferences Can Change Top Large Language Model Rankings

ResearchDGX agent

arXiv:2508.11847v4 Announce Type: replace-cross Abstract: We propose a method for evaluating the robustness of widely used LLM ranking systems -- variants of a Bradley--Terry model -- to dropping a wo

Game Theory Driven Multi-Agent Framework Mitigates Language Model Hallucination

AgentsDGX agent

arXiv:2607.08403v1 Announce Type: new Abstract: The application of lightweight Large Language Models in rule-based scientific domains remains severely limited by their tendency to mimic linguistic pat

Large-Language-Models-as-a-Judge in Theory-Agnostic Adaptive Metric-Alignment for Prototypical Networks in Personality Recognition

SafetyDGX agent

arXiv:2607.08374v1 Announce Type: cross Abstract: Personality recognition has traditionally been constrained by theory-dependent formulations, where models are trained to fit predefined psychological

Many people were doubting Meta's position in the AI race. Yesterday, they dropped Muse Spark 1.1, now one of the strongest agentic models, a…

AgentsDGX agent

Many people were doubting Meta's position in the AI race. Yesterday, they dropped Muse Spark 1.1, now one of the strongest agentic models, and massively undercut OpenAI and Anthropic on price. When I

Omni-Sleep: A Sleep Foundation Model via Hierarchical Contrastive Learning of CNS--ANS Dynamic

ResearchDGX agent

arXiv:2607.07720v1 Announce Type: cross Abstract: Sleep physiology arises from the coordinated dynamics of the central nervous system (CNS) and autonomic nervous system (ANS), as reflected by multimod

Structural Bottlenecks on Frequency Representation in End-to-End Audio Models

ResearchDGX agent

arXiv:2607.08545v1 Announce Type: cross Abstract: End-to-end neural audio models achieve high-fidelity compression and generation. We might read that performance as evidence they directly represent in

Today, together with @kyutai_labs, we’re introducing our new Audio-to-MIDI model. It takes a finished recording, identifies the instruments …

ResearchDGX agent

Today, together with @kyutai_labs, we’re introducing our new Audio-to-MIDI model. It takes a finished recording, identifies the instruments playing, and returns separate MIDI tracks for each — voice,

Towards Isolated Interventions via Almost Orthogonal Features in Language Models

Local AiDGX agent

arXiv:2602.04718v2 Announce Type: replace-cross Abstract: A central premise in mechanistic interpretability is that meaningful concepts in language models are represented by linear features in activat

TypeProbe: Recovering Type Representations from Hidden States of Pre-trained Code Models

ResearchDGX agent

arXiv:2607.08339v1 Announce Type: cross Abstract: State-of-the-art code models achieve impressive performance, yet the extent to which they internally encode type information remains poorly understood

XOV-Action: Towards Generalizable Open-Vocabulary Action Recognition

Model ReleasesDGX agent

arXiv:2403.01560v3 Announce Type: replace Abstract: Inspired by the impressive success of image-text foundation models, recent works have proposed to adapt these foundation models to video data, leadi

9 Jul 2026

Big day for Ollama! When we started, open models and the open source AI ecosystem were in their early days with few believers. Our belief in…

Local AiDGX agent

Big day for Ollama! When we started, open models and the open source AI ecosystem were in their early days with few believers. Our belief in open source has never wavered. With today's fundraising ann

CoFrGeNets replace the ‘bones’ of transformer-based models

ResearchDGX agent

CoFrGeNets are a novel architecture proposed by IBM Research that replaces the core structural components ('bones') of transformer-based models with a more efficient design. This approach aims to impr

Congrats to @ollama on the fundraise and 9M+ active builders.🚀 Open models are becoming the default path for developers who want to build, …

Local AiDGX agent

Congrats to @ollama on the fundraise and 9M+ active builders.🚀 Open models are becoming the default path for developers who want to build, run, and own their AI stack. We're excited to support the eco

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models

SafetyDGX agent

arXiv:2602.23802v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have shown remarkable progress in visual reasoning and understanding tasks but still struggle to capture th

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety

Model ReleasesDGX agent

arXiv:2607.07695v1 Announce Type: new Abstract: We introduce institutional red-teaming, an evaluation methodology for testing deployment rules in multi-agent AI: hold the agents, objectives, and task

See how every model compares: http://cursor.com/evals

ToolsDGX agent

Cursor shared a link to their evals page that provides a comprehensive comparison of different AI models' performance and capabilities. The page likely displays benchmarks, metrics, or evaluation resu

Strategies for Span Labeling with Large Language Models

ResearchDGX agent

arXiv:2601.16946v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for text analysis tasks, such as named entity recognition or error detection. Unlike encoder-base

Vision Foundation Models in Radiology: A Scoping Review of Data, Methodology, Evaluation and Clinical Translation

SafetyDGX agent

arXiv:2607.07219v1 Announce Type: cross Abstract: Vision foundation models (VFMs) are increasingly being developed for radiological imaging, yet their definition, development and evaluation remain het

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time

ResearchDGX agent

arXiv:2607.06988v1 Announce Type: cross Abstract: Steering robot foundation models (RFMs) toward new task variants or user-preferred behaviors remains challenging, often requiring additional robot dem

8 Jul 2026

A Coin Flip Per Token: Bernoulli Sparse Steering of Large Language Models

SafetyDGX agent

arXiv:2607.05615v1 Announce Type: new Abstract: Activation steering via sparse autoencoders (SAEs) enables behavioral control of large language models without task-specific fine-tuning, but standard m

AVA-VLM: Adaptive Visual Attention-Vision Language Model for In-the-Wild Construction Site Monitoring

Local AiDGX agent

arXiv:2607.05859v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are promising for construction-site monitoring, and recent construction-tailored VLMs have primarily adapted pretrained VL

Diffusion Language Model Parallel Decoding via Product-of-Experts Bridge

ResearchDGX agent

arXiv:2606.08048v1 Announce Type: cross Abstract: Diffusion language models (DLMs) offer substantial speed advantages through parallel decoding, but the lack of token dependencies limits generation qu

From Global to Granular: Revealing IQA Model Performance via Correlation Surface

ResearchDGX agent

arXiv:2601.21738v2 Announce Type: replace-cross Abstract: Evaluation of Image Quality Assessment (IQA) models has long been dominated by global correlation metrics, such as Pearson Linear Correlation

Full-range Binary Classifier Calibration for Stable Model Updates in Production

ApplicationsDGX agent

arXiv:2607.05481v1 Announce Type: cross Abstract: Detection models running in adversarial environments face a malicious distribution that drifts rapidly while the benign distribution stays comparative

Hypothesis-driven Model Expansion under Uncertainty for Open-World Robot Planning

AgentsDGX agent

arXiv:2607.06501v1 Announce Type: new Abstract: We consider an open-world planning setting in which service robots must operate in unknown environments with incomplete knowledge of objects and actions

Life Cycle Assessment of Pre-training the Lucie 7B Open-Source Large Language Model on the Jean Zay Supercomputer

HardwareDGX agent

arXiv:2607.05408v1 Announce Type: cross Abstract: The environmental impact of training large language models (LLMs) is increasingly scrutinised, yet most published estimates focus on operational energ

MLLM-LLaVA-FL: Multimodal Large Language Model Assisted Federated Learning

Local AiDGX agent

arXiv:2409.06067v3 Announce Type: replace Abstract: Previous studies on federated learning (FL) often encounter performance degradation due to data heterogeneity among different clients. In light of t

most agent labs are shy about acknowledging chinese model use because they need to sell to gov/defense cog team did the hard part to product…

AgentsDGX agent

most agent labs are shy about acknowledging chinese model use because they need to sell to gov/defense cog team did the hard part to productionize: 1. build a multilingual propaganda & censorship eval

Narrative World Model: Narratology-Grounded Writer Memory for Long-Form Fiction

Model ReleasesDGX agent

arXiv:2607.05577v1 Announce Type: new Abstract: Long-form fiction writers need memory that answers multi-hop questions about evolving story state: who knows a secret and when they learned it, whether

Native-speed vLLM transformers modeling backend

ToolsDGX agent

This entry covers the integration of vLLM, a high-performance inference engine, with Hugging Face's transformers library to enable faster language model inference. The native-speed backend allows user

Reward as An Agent for Embodied World Models

AgentsDGX agent

arXiv:2606.19990v2 Announce Type: replace Abstract: While RL has become a promising tool for refining world models, existing methods largely rely on conservative rollouts near the training distributio

Social 3D Scene Graphs: Modeling Human Actions and Relations for Interactive Service Robots

Model ReleasesDGX agent

arXiv:2509.24966v2 Announce Type: replace Abstract: Understanding how people interact with their surroundings and each other is essential for enabling robots to act in socially compliant and context-a

Structured Data Extraction from Real Estate Documents using Clustering, Classification, and Large Language Models

Model ReleasesDGX agent

arXiv:2607.06012v1 Announce Type: new Abstract: Real estate property listings expose structured metadata through the API. Still, the richest property-level information (i.e., legal status, structural

Training-Free Acceleration for Vision-Language-Action Models with Action Caching and Refinement

ApplicationsDGX agent

arXiv:2607.06370v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising approach for generalizable robotic manipulations. In particular, flow matching-based V

7 Jul 2026

A harmonised dataset for Earth system foundation models

TutorialsDGX agent

arXiv:2607.03298v1 Announce Type: cross Abstract: Foundation models for Earth systems have so far been trained primarily on physical climate and weather data, with limited representation of the human

AgentGym2: Benchmarking Large Language Model Agents in De-Idealized Real-World Environments

Model ReleasesDGX agent

arXiv:2607.05174v1 Announce Type: new Abstract: Language agents, i.e., LLM agents, progress rapidly and are increasingly deployed in production environments. This trend underscores the urgent need for

AMT-APC: Automatic Piano Cover by Fine-Tuning an Automatic Music Transcription Model

ResearchDGX agent

arXiv:2409.14086v2 Announce Type: replace-cross Abstract: There have been several studies on automatically generating piano covers, and recent advancements in deep learning have enabled the creation o

APeB: Benchmarking Personalization Ability of Large Language Model Agents

Model ReleasesDGX agent

arXiv:2607.03162v1 Announce Type: new Abstract: LLM-powered agents struggle with personalization when users issue raw, underspecified queries. In this setting, agents must infer latent intent, extract

AquaGen: Scaling generative models to molecular dynamics precision on thousands of atoms

HardwareDGX agent

arXiv:2607.03513v1 Announce Type: cross Abstract: We present AquaGen, the first all-atom, explicit solvent, periodic-boundary-condition-aware generative model that produces molecular configurations fr

CanniUplift: A Holistic Framework for Mitigating Seller and Incentive Cannibalization in E-commerce Uplift Modeling

SafetyDGX agent

arXiv:2607.05242v1 Announce Type: cross Abstract: Personalized incentive allocation is vital for e-commerce, where uplift modeling is the standard for estimating Individual Treatment Effects (ITE). Ho

EyeMulator: Improving Code Language Models by Mimicking Human Visual Attention

TutorialsDGX agent

arXiv:2508.16771v3 Announce Type: replace-cross Abstract: Code Language Models (CodeLLMs) learn token importance from data correlations, whereas human developers attend selectively to semantically sal

Gemma 4 Technical Report

Model ReleasesDGX agent

arXiv:2607.02770v1 Announce Type: cross Abstract: We introduce Gemma 4, a new generation of open-weight, natively multimodal language models in the Gemma model family. Designed to advance compute effi

GeoFlow: Geo-Aware Modeling of Inter-Area Relationships in Origin-Destination Flow Prediction and Generation

ResearchDGX agent

arXiv:2607.05257v1 Announce Type: new Abstract: Origin-destination (OD) flow modeling underpins urban planning and mobility analysis, but prevailing graph-based methods often neglect salient geographi

Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling

ResearchDGX agent

arXiv:2607.02980v1 Announce Type: cross Abstract: Scaling modern large language models (LLMs) to long contexts is limited by the quadratic computation cost, and poor length extrapolation of dense atte

← Previous
1…150151152153154…1010
Next →