AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
4 Jun 2026

Large Language Models in K-12 Education: Alignment with State Curriculum Standards and Student Personas

SafetyDGX agent

arXiv:2606.04846v1 Announce Type: new Abstract: As Large Language Models (LLMs) become increasingly popular in educational settings, they raise important questions about the ethical implications of th

Model Evaluations: Prove Your Routing Policy Actually Works

SafetyDGX agent

This article discusses methods and tools for evaluating routing policies in machine learning models, likely covering techniques to validate that model routing decisions are effective and functioning a

Model-Preserving Adaptive Rounding

ResearchDGX agent

arXiv:2505.22988v3 Announce Type: replace-cross Abstract: The goal of quantization is to produce a compressed model whose output distribution is as close to the original model's as possible. To do thi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.04773v1 Announce Type: cross Abstract: Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffe

Reconciling Causality and Non-Equilibrium Thermodynamics with Hamiltonian Causal Models

ApplicationsDGX agent

arXiv:2606.04822v1 Announce Type: new Abstract: Causal modeling of physical temporal phenomena must handle interventions that act along trajectories, nonstationary induced laws, path-dependent effects

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

Local AiDGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficie…

ApplicationsDGX agent

reiterating: 'We're using the more expensive models to explore. Once we scale some of these experiences, we'll look to bring in more efficient models that are more efficient on a token basis or are op

Selecting haptic guidance models in teleoperation: guidelines from a comparative user study

ResearchDGX agent

arXiv:2606.04157v1 Announce Type: new Abstract: Haptic guidance in teleoperation enhances operator performance through force feedback. This paper presents guidelines to select the most appropriate mod

SMAC-Talk: A Natural Language Extension of the StarCraft Multi-Agent Challenge for Large Language Models

Model ReleasesDGX agent

arXiv:2606.04202v1 Announce Type: new Abstract: As LLMs become more widely deployed, they are increasingly expected to work alongside other AI agents rather than operating in isolation. Effective coor

Toward Trustworthy Portrait Editing: Evaluation of Demographic Misrepresentation in I2I Models

Model ReleasesDGX agent

arXiv:2602.16149v2 Announce Type: replace Abstract: Instruction-guided image-to-image (I2I) editors are increasingly used in consumer and professional visual workflows, where trustworthiness depends n

Validity Threats for Foundation Model Research

ResearchDGX agent

arXiv:2606.05029v1 Announce Type: cross Abstract: Controlled experiments are the backbone of machine learning research, but at the scale of modern foundation models, they have become prohibitively exp

When Offline Selectors Cannot Beat the Best Single Model: A Diagnostic Study on edX Dropout Prediction

ResearchDGX agent

arXiv:2606.04161v1 Announce Type: new Abstract: Different predictors often excel on different inputs, so picking the best one per instance promises higher accuracy than committing to a single model. I

3 Jun 2026

Brief Announcement: Generative Markov Model for Distributed Computing Systems

SafetyDGX agent

arXiv:2606.03061v1 Announce Type: cross Abstract: Emerging distributed computing paradigms, such as the computing continuum, are inherently heterogeneous, stochastic, and complex. Efficiently and effe

CANMOT: Class-Aware Noise Modeling for Multi-Object Tracking in Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.03590v1 Announce Type: new Abstract: Kalman filter (KF)-based multi-object tracking (MOT) remains a strong baseline for autonomous driving due to its strong performance, computational effic

Clustered Self-Assessment: A Simple yet Effective Method for Uncertainty Quantification in Large Language Models

ResearchDGX agent

arXiv:2606.03846v1 Announce Type: cross Abstract: Large language models (LLMs) demonstrate remarkable performance across diverse tasks, but they often generate responses that appear plausible while be

Fresh Open weight model drop😋

Local AiDGX agent

Fresh Open weight model drop😋 Introducing Ideogram 4.0: the best open image model in the world. Think it. Make it. Own it. Download the weights, fine-tune on your own data, and run it on your hardware

Fully Automated Identification of Lexical Alignment and Preference-Stage Shifts in Large Language Models

Model ReleasesDGX agent

arXiv:2606.03165v1 Announce Type: cross Abstract: The language used by digital chat assistants such as ChatGPT can diverge from human expectations (misalignment). Research, mostly on Scientific Englis

Linear Probes Detect Task Format, Not Reasoning Mode in Language Model Hidden States

TutorialsDGX agent

arXiv:2606.02907v1 Announce Type: cross Abstract: Linear probing of large language model (LLM) hidden states is widely used to claim that models learn distinct representations for different reasoning

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

SafetyDGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

Multiple Choice Learning of Low-Rank Adapters for Language Modeling

ResearchDGX agent

arXiv:2507.10419v3 Announce Type: replace-cross Abstract: We propose LoRA-MCL, a training scheme that extends next-token prediction in language models with a method designed to decode diverse, plausib

Neural Fields as World Models

Local AiDGX agent

arXiv:2602.18690v2 Announce Type: replace-cross Abstract: Humans rehearse possible futures offline, as in mental practice and perhaps dreaming, suggesting that world models may support task learning a

PHASER: Phase-Aware and Semantic Experience Replay for Vision-Language-Action Models

AgentsDGX agent

arXiv:2606.03598v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved remarkable success in language-conditioned robotic manipulation. However, deploying these models in

PRISM: Synergizing Vision Foundation Models via Self-organized Expert Specialization

ResearchDGX agent

arXiv:2606.03444v1 Announce Type: cross Abstract: Unifying the complementary strengths of diverse Vision Foundation Models (VFMs) into a single efficient model is highly desirable but challenged by th

Quantifying Faithful Confidence Expression in Large Reasoning Models

SafetyDGX agent

arXiv:2606.03969v1 Announce Type: cross Abstract: Reliable uncertainty communication is critical to the trustworthiness of LLMs, yet faithful calibration (FC)--the alignment between models' intrinsic

Releasing vui an open source voice mode 300M TTS model Runs on a single consumer gpu / apple sillicon Context aware speech 6 minutes of cont…

HardwareDGX agent

Jeremy Howard announced the release of Vui, an open-source voice mode text-to-speech (TTS) model with 300 million parameters that can run on consumer GPUs and Apple Silicon. The model features context

Routing and post-training open-source models won't only give you more accurate systems but also meaningfully faster and cheaper systems as m…

ApplicationsDGX agent

Routing and post-training open-source models won't only give you more accurate systems but also meaningfully faster and cheaper systems as most companies are currently learning (in addition to giving

Scalable Single-Cell Gene Expression Generation with Latent Diffusion Models

ResearchDGX agent

arXiv:2511.02986v2 Announce Type: replace-cross Abstract: Computational modeling of single-cell gene expression is crucial for understanding cellular processes, but generating realistic expression pro

Selective Token-Level Cryptographic Redaction for Privacy-Preserving Clinical Deployment of Large Language Models

SafetyDGX agent

arXiv:2606.03399v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly used for clinical applications, many existing pipelines require sending raw sensitive health informa

SketchSong: Hierarchical Song Generation with Sketch Planning and Fine-Grained Multi-Track Modeling

ResearchDGX agent

arXiv:2606.03169v1 Announce Type: cross Abstract: Recent song generation systems can synthesize realistic audio, yet generating complete songs remains challenging for two reasons. First, explicit song

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on ho…

TutorialsDGX agent

the best agents aren't just built with the best models: they're built with harnesses purpose-built for the task at hand here's a guide on how to build a harness that's really good at feeding the model

theUSshould lead on AI by continuing to develop the very best models, making sure they're safe, and getting cyber tools into the hands of tr…

IndustryDGX agent

Sam Altman argues that US leadership in artificial intelligence requires three concurrent priorities: advancing cutting-edge AI model development, ensuring these models incorporate robust safety measu

Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models

ApplicationsDGX agent

arXiv:2606.03748v1 Announce Type: cross Abstract: Real-time vision demands models that are accurate, efficient, and simple to deploy across diverse hardware. The YOLO family has become widely deployed

Value-Aware Stochastic KV Cache Eviction for Reasoning Models

ResearchDGX agent

arXiv:2606.03928v1 Announce Type: cross Abstract: Reasoning models improve accuracy through extended chains of thought, but their long outputs create a memory and compute bottleneck. KV cache eviction

2 Jun 2026

A Lightweight Deep Learning-based Model for Ranking Influential Nodes in Complex Networks

Local AiDGX agent

arXiv:2507.19702v1 Announce Type: cross Abstract: Identifying influential nodes in complex networks is a critical task with a wide range of applications across different domains. However, existing app

An Algebraic View of the Expressivity of Recurrent Language Models

ApplicationsDGX agent

arXiv:2606.01765v1 Announce Type: cross Abstract: What formal languages can a recurrent neural language model recognize? Formal results in the literature conflict: some authors report Turing-completen

Bayesian meta-learning for modeling Alzheimer's disease progression

ApplicationsDGX agent

arXiv:2606.02228v1 Announce Type: cross Abstract: Predicting whether an individual with Alzheimer's disease will experience mild or severe disease progression is essential for personalized treatment.

BERT4beam: Large AI Model Enabled Generalized Beamforming Optimization

ResearchDGX agent

arXiv:2509.11056v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) is anticipated to emerge as a pivotal enabler for the forthcoming sixth-generation (6G) wireless communication sy

Breaking the Reversal Curse in Autoregressive Language Models via Identity Bridge

SafetyDGX agent

arXiv:2602.02470v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) have achieved remarkable success in many complex tasks, yet they can still fail in very simple logical r

CARTE: A Benchmark for Mapping Language Model Knowledge Across France

Model ReleasesDGX agent

arXiv:2606.01995v1 Announce Type: new Abstract: We introduce CARTE 1 (Culturally Anchored Regional-Territorial Evaluation), a multiplechoice benchmark for evaluating the ability of large language mode

d2: Improving Reasoning in Diffusion Language Models via Trajectory Likelihood Estimation

SafetyDGX agent

arXiv:2509.21474v4 Announce Type: replace Abstract: While diffusion language models (DLMs) have achieved competitive performance in text generation, improving their reasoning ability with reinforcemen

Decoding in Order-Agnostic Language Models: Chain-Rule Deviation and Uniform Spreading

ResearchDGX agent

arXiv:2606.00997v1 Announce Type: new Abstract: Order-agnostic language models (OALMs), including discrete diffusion language models (dLLMs), are trained to predict masked tokens under arbitrary condi

Efficient Test-time Inference for Generative Planning Models

ResearchDGX agent

arXiv:2606.00618v1 Announce Type: new Abstract: Generative models have emerged as a powerful paradigm for AI planning, yet their performance remains constrained by the training data distribution. One

Error Bounds for a Diffusion Model-Based Drift Estimator

Model ReleasesDGX agent

arXiv:2606.02115v1 Announce Type: cross Abstract: Parameter estimation in stochastic differential equations is a classical statistical problem of much importance in many scientific fields. Recent work

Evaluating Real-World Generalizability of Algorithm Selection Models

Model ReleasesDGX agent

arXiv:2606.02016v1 Announce Type: new Abstract: Algorithm Selection (AS) aims to automatically identify the most suitable optimization algorithm for a given problem instance by leveraging measurable p

Evaluating the Performance of Deep Learning Models in Whole-body Dynamic 3D Posture Prediction During Load-reaching Activities

ResearchDGX agent

arXiv:2511.20615v2 Announce Type: replace-cross Abstract: This study aimed to explore the application of deep neural networks for whole-body human posture prediction during dynamic load-reaching activ

Exploring the Capabilities of Large Language Model Encoders for Image-Text Retrieval in Chest X-rays

Model ReleasesDGX agent

arXiv:2509.15234v2 Announce Type: replace Abstract: Multimodal learning from paired medical images and clinical text is a central challenge in medical data-driven informatics, where effective cross-mo

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

HardwareDGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

Failure of contextual invariance in large language models

SafetyDGX agent

arXiv:2603.23485v2 Announce Type: replace-cross Abstract: Standard evaluation practices assume that large language model (LLM) outputs are stable when prompts are embedded in contextually equivalent d

FATE-VLA:Failue-aware test generation for vision-language-action models

ResearchDGX agent

arXiv:2606.02307v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are increasingly used as generalist robot policies, yet their evaluation still relies largely on static benchmarks t

Fine-Tuning Diffusion Models for Molecular Generation via Reinforcement Learning and Fast Sampling

Model ReleasesDGX agent

arXiv:2606.01220v1 Announce Type: cross Abstract: Generating molecules that simultaneously satisfy drug-like properties and conform to the 3D structure of a target protein is a core challenge in struc

Fine-Tuning Without Forgetting In-Context Learning: A Theoretical Analysis of Linear Attention Models

ApplicationsDGX agent

arXiv:2602.23197v2 Announce Type: replace Abstract: Transformer-based large language models exhibit in-context learning, enabling adaptation to downstream tasks via few-shot prompting with demonstrati

FLARE: Diffusion for Hybrid Language Model

HardwareDGX agent

arXiv:2606.01774v1 Announce Type: cross Abstract: Autoregressive (AR) large language models (LLMs) have achieved broad practical success, but sequential decoding remains a key bottleneck for low-laten

From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models

SafetyDGX agent

arXiv:2606.00083v1 Announce Type: cross Abstract: Reinforcement learning relies on accurate reward functions, which are often hand-crafted or even unavailable in real-world applications, such as robot

How AI Fails: An Interactive Pedagogical Tool for Demonstrating Dialectal Bias in Automated Toxicity Models

Model ReleasesDGX agent

arXiv:2511.06676v3 Announce Type: replace Abstract: Now that AI-driven moderation has become pervasive in everyday life, we often hear claims that 'the AI is biased'. While this is often said jokingly

HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single Image

ResearchDGX agent

arXiv:2606.02573v1 Announce Type: new Abstract: In this paper, we present HumanNOVA, a photorealistic, universal, and rapid model for generating 3D human avatars from a single RGB image. Achieving bot

Intercepting the Future: Latent-Space Predictive World Model for Dynamic VLA Manipulation

ResearchDGX agent

arXiv:2606.02486v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models generalize across static manipulation but fail when objects move during task execution. They map the current observa

Lessons from the Trenches on Reproducible Evaluation of Language Models

ResearchDGX agent

arXiv:2405.14782v3 Announce Type: replace Abstract: Reliable evaluation of language models (LMs) remains an open challenge. Re- searchers and engineers face methodological issues such as the sensitivi

Limits of Spatial Imagery Reasoning in Frontier LLM Models

ResearchDGX agent

arXiv:2603.26779v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated impressive reasoning capabilities, yet they struggle with spatial tasks that require mental sim

Physics-Encoded Inverse Modeling for Arctic Snow Depth Prediction

Model ReleasesDGX agent

arXiv:2601.17074v4 Announce Type: replace-cross Abstract: Accurate estimation in time-varying inverse problems under limited and sparse observations remains a fundamental challenge across scientific d

Physics-Informed Modeling and Control of Emergent Behaviors in Robot Swarms

Local AiDGX agent

arXiv:2606.01597v1 Announce Type: new Abstract: Robot swarms can exhibit coherent collective behaviors through local perception, limited communication and decentralized decision-making, yet modeling a

← Previous
1…132133134135136…1009
Next →