AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,055 results
Model Releases

SOLAR: Communication-Efficient Model Adaptation via Subspace-Oriented Latent Adapter Reparametrization

DGX agent

arXiv:2604.08368v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) methods, such as LoRA, enable scalable adaptation of foundation models by injecting low-rank adapters. However,

model-releasesarxiv-cs-cl
10 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SpatialMosaic: A Multiview VLM Dataset for Partial Visibility

DGX agent

arXiv:2512.23365v3 Announce Type: replace Abstract: The rapid progress of Multimodal Large Language Models (MLLMs) has unlocked the potential for enhanced 3D scene understanding and spatial reasoning.

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Spatio-Temporal Grounding of Large Language Models from Perception Streams

DGX agent

arXiv:2604.07592v1 Announce Type: new Abstract: Embodied-AI agents must reason about how objects move and interact in 3-D space over time, yet existing smaller frontier Large Language Models (LLMs) st

model-releasesarxiv-cs-ro
10 Apr 2026
Model Releases

SpecQuant: Spectral Decomposition and Adaptive Truncation for Ultra-Low-Bit LLMs Quantization

DGX agent

arXiv:2511.11663v2 Announce Type: replace-cross Abstract: The emergence of accurate open large language models (LLMs) has sparked a push for advanced quantization techniques to enable efficient deploy

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Spectral Edge Dynamics Reveal Functional Modes of Learning

DGX agent

arXiv:2604.06256v1 Announce Type: cross Abstract: Training dynamics during grokking concentrate along a small number of dominant update directions -- the spectral edge -- which reliably distinguishes

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Stacked from One: Multi-Scale Self-Injection for Context Window Extension

DGX agent

arXiv:2603.04759v2 Announce Type: replace Abstract: The limited context window of contemporary large language models (LLMs) remains a primary bottleneck for their broader application across diverse do

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Stitch4D: Sparse Multi-Location 4D Urban Reconstruction via Spatio-Temporal Interpolation

DGX agent

arXiv:2604.07923v1 Announce Type: new Abstract: Dynamic urban environments are often captured by cameras placed at spatially separated locations with little or no view overlap. However, most existing

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Stochastic Auto-conditioned Fast Gradient Methods with Optimal Rates

DGX agent

arXiv:2604.06525v1 Announce Type: cross Abstract: Achieving optimal rates for stochastic composite convex optimization without prior knowledge of problem parameters remains a central challenge. In the

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation

DGX agent

arXiv:2604.07028v1 Announce Type: cross Abstract: Strategic interaction in adversarial domains such as law, diplomacy, and negotiation is mediated by language, yet most game-theoretic models abstract

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses

DGX agent

arXiv:2602.22683v2 Announce Type: replace Abstract: The rapid advancement of AI-powered smart glasses-one of the hottest wearable devices-has unlocked new frontiers for multimodal interaction, with Vi

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding

DGX agent

arXiv:2604.07753v1 Announce Type: cross Abstract: Empowering Large Multimodal Models (LMMs) with image generation often leads to catastrophic forgetting in understanding tasks due to severe gradient c

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Synthetic Homes: A Multimodal Generative AI Pipeline for Residential Building Data Generation under Data Scarcity

DGX agent

arXiv:2509.09794v4 Announce Type: replace Abstract: Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

T-Gated Adapter: A Lightweight Temporal Adapter for Vision-Language Medical Segmentation

DGX agent

arXiv:2604.08167v1 Announce Type: new Abstract: Medical image segmentation traditionally relies on fully supervised 3D architectures that demand a large amount of dense, voxel-level annotations from c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Tabular GANs for uneven distribution

DGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models

DGX agent

arXiv:2604.06291v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of Large Language Models (LLMs), and recent Mixture-of-Experts (MoE) extensions fur

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TeamLLM: A Human-Like Team-Oriented Collaboration Framework for Multi-Step Contextualized Tasks

DGX agent

arXiv:2604.06765v1 Announce Type: cross Abstract: Recently, multi-Large Language Model (LLM) frameworks have been proposed to solve contextualized tasks. However, these frameworks do not explicitly em

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TEMPER: Testing Emotional Perturbation in Quantitative Reasoning

DGX agent

arXiv:2604.07801v1 Announce Type: new Abstract: Large language models are trained and evaluated on quantitative reasoning tasks written in clean, emotionally neutral language. However, real-world quer

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Temporally Phenotyping GLP-1RA Case Reports with Large Language Models: A Textual Time Series Corpus and Risk Modeling

DGX agent

arXiv:2604.06197v1 Announce Type: cross Abstract: Type 2 diabetes case reports describe complex clinical courses, but their timelines are often expressed in language that is difficult to reuse in long

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Tensor-Augmented Convolutional Neural Networks: Enhancing Expressivity with Generic Tensor Kernels

DGX agent

arXiv:2604.08072v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) excel at extracting local features hierarchically, but their performance in capturing complex correlations hinges h

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Tensor-Efficient High-Dimensional Q-learning

DGX agent

arXiv:2511.03595v2 Announce Type: replace Abstract: High-dimensional reinforcement learning(RL) faces challenges with complex calculations and low sample efficiency in large state-action spaces. Q-lea

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you …

DGX agent

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you to retrieve context from your private notebooks or convert y

model-releasesgoogle-ai--x
10 Apr 2026
Model Releases

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

DGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The ATOM Report: Measuring the Open Language Model Ecosystem

DGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning

DGX agent

arXiv:2604.06427v1 Announce Type: cross Abstract: The viability of chain-of-thought (CoT) monitoring hinges on models being unable to reason effectively in their latent representations. Yet little is

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Geometry of Forgetting

DGX agent

arXiv:2604.06222v1 Announce Type: cross Abstract: Why do we forget? Why do we remember things that never happened? The conventional answer points to biological hardware. We propose a different one: ge

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Impact of Steering Large Language Models with Persona Vectors in Educational Applications

DGX agent

arXiv:2604.07102v1 Announce Type: cross Abstract: Activation-based steering can personalize large language models at inference time, but its effects in educational settings remain unclear. We study pe

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Stepwise Informativeness Assumption: Why are Entropy Dynamics and Reasoning Correlated in LLMs?

DGX agent

arXiv:2604.06192v1 Announce Type: cross Abstract: Recent work uses entropy-based signals at multiple representation levels to study reasoning in large language models, but the field remains largely em

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Theory and Practice of Highly Scalable Gaussian Process Regression with Nearest Neighbours

DGX agent

arXiv:2604.07267v1 Announce Type: cross Abstract: Gaussian process (GP) regression is a widely used non-parametric modeling tool, but its cubic complexity in the training size limits its use on mass

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

The Traveling Thief Problem with Time Windows: Benchmarks and Heuristics

DGX agent

arXiv:2604.06724v1 Announce Type: cross Abstract: While traditional optimization problems were often studied in isolation, many real-world problems today require interdependence among multiple optimiz

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

There were some exceptionally cool demos from @ollama and omlx using MLX to run Qwen 3.5 and Gemma 4 on Apple silicon. The capabilities of l…

DGX agent

There were some exceptionally cool demos from @ollama and omlx using MLX to run Qwen 3.5 and Gemma 4 on Apple silicon. The capabilities of local LLMs and the surrounding ecosystem have come a long way

model-releasesollama--x
10 Apr 2026
Model Releases

Thompson Sampling for Infinite-Horizon Discounted Decision Processes

DGX agent

arXiv:2405.08253v3 Announce Type: replace-cross Abstract: This paper develops a viable notion of learning for sampling-based algorithms that applies in broader settings than previously considered. Mor

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Tight Convergence Rates for Online Distributed Linear Estimation with Adversarial Measurements

DGX agent

arXiv:2604.06282v1 Announce Type: cross Abstract: We study mean estimation of a random vector X in a distributed parameter-server-worker setup. Worker i observes samples of a_i^op X, where $a_

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model

DGX agent

arXiv:2604.07816v1 Announce Type: new Abstract: Tool learning has emerged as a promising paradigm for large language models (LLMs) to address real-world challenges. Due to the extensive and irregularl

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Toward a universal foundation model for graph-structured data

DGX agent

arXiv:2604.06391v1 Announce Type: cross Abstract: Graphs are a central representation in biomedical research, capturing molecular interaction networks, gene regulatory circuits, cell--cell communicati

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

DGX agent

arXiv:2505.22976v2 Announce Type: replace-cross Abstract: The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Towards Accurate and Calibrated Classification: Regularizing Cross-Entropy From A Generative Perspective

DGX agent

arXiv:2604.06689v1 Announce Type: new Abstract: Accurate classification requires not only high predictive accuracy but also well-calibrated confidence estimates. Yet, modern deep neural networks (DNNs

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval

DGX agent

arXiv:2512.08410v2 Announce Type: replace Abstract: Due to excessive memory overhead, most Multimodal Large Language Models (MLLMs) can only process videos of limited frames. In this paper, we propose

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

DGX agent

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

DGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TR-EduVSum: A Turkish-Focused Dataset and Consensus Framework for Educational Video Summarization

DGX agent

arXiv:2604.07553v1 Announce Type: new Abstract: This study presents a framework for generating the gold-standard summary fully automatically and reproducibly based on multiple human summaries of Turki

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

DGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Training Data Size Sensitivity in Unsupervised Rhyme Recognition

DGX agent

arXiv:2604.08156v1 Announce Type: new Abstract: Rhyme is deceptively intuitive: what is or is not a rhyme is constructed historically, scholars struggle with rhyme classification, and people disagree

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

DGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Ultraplan uses roughly the same number of tokens (and subscription rate limits) as plan mode. See the docs for more: http://docs.claude.com/…

DGX agent

Ultraplan, a planning feature in Claude, consumes approximately the same number of tokens and counts against subscription rate limits similarly to standard plan mode. Users should be aware that using

model-releasesthariq--x
10 Apr 2026
Model Releases

Unifying Speech Editing Detection and Content Localization via Prior-Enhanced Audio LLMs

DGX agent

arXiv:2601.21463v2 Announce Type: replace-cross Abstract: Existing speech editing detection (SED) datasets are predominantly constructed using manual splicing or limited editing operations, resulting

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

UniversalVTG: A Universal and Lightweight Foundation Model for Video Temporal Grounding

DGX agent

arXiv:2604.08522v1 Announce Type: new Abstract: Video temporal grounding (VTG) is typically tackled with dataset-specific models that transfer poorly across domains and query styles. Recent efforts to

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Validated Intent Compilation for Constrained Routing in LEO Mega-Constellations

DGX agent

arXiv:2604.07264v1 Announce Type: cross Abstract: Operating LEO mega-constellations requires translating high-level operator intents ('reroute financial traffic away from polar links under 80 ms') int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents

DGX agent

arXiv:2603.15118v2 Announce Type: replace Abstract: We introduce VAREX (VARied-schema EXtraction), a benchmark for evaluating multimodal foundation models on structured data extraction from government

model-releasesarxiv-cs-cv
10 Apr 2026
← Previous
1…454455456457458…460
Next →