AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
Research

MultiTok: Variable-Length Tokenization for Efficient LLMs Adapted from LZW Compression

DGX agent

arXiv:2410.21548v3 Announce Type: replace Abstract: Large language models have drastically changed the prospects of AI by introducing technologies for more complex natural language processing. However

researcharxiv-cs-cl
27 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Neural Recovery of Historical Lexical Structure in Bantu Languages from Modern Data

DGX agent

arXiv:2604.22730v1 Announce Type: cross Abstract: We investigate whether neural models trained exclusively on modern morphological data can recover cross-lingual lexical structure consistent with hist

researcharxiv-cs-cl
27 Apr 2026
Research

Non-Minimal Sampling and Consensus for Prohibitively Large Datasets

DGX agent

arXiv:2604.22518v1 Announce Type: new Abstract: We introduce NONSAC (Non-Minimal Sampling and Consensus), a general framework for robust and scalable model estimation from arbitrarily large datasets c

researcharxiv-cs-cv
27 Apr 2026
Model Releases

Parameter-Efficient Conditioning for Material Generalization in Graph-Based Simulators

DGX agent

arXiv:2511.05456v2 Announce Type: replace Abstract: Graph network-based simulators (GNS) have demonstrated strong potential for learning particle-based physics (such as fluids, deformable solids, and

model-releasesarxiv-cs-lg
27 Apr 2026
Model Releases

PSI: A Benchmark for Human Interpretation and Response in Traffic Interactions

DGX agent

arXiv:2112.02604v3 Announce Type: replace-cross Abstract: Accurately modeling pedestrian intention and understanding driver decision-making processes are critical for the development of safe and socia

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Regularized Meta-Learning for Improved Generalization

DGX agent

arXiv:2602.12469v2 Announce Type: replace Abstract: Deep ensemble methods often improve predictive performance, yet they suffer from three practical limitations: redundancy among base models that infl

model-releasesarxiv-cs-lg
27 Apr 2026
Research

Removing Sandbagging in LLMs by Training with Weak Supervision

DGX agent

arXiv:2604.22082v1 Announce Type: cross Abstract: As AI systems begin to automate complex tasks, supervision increasingly relies on weaker models or limited human oversight that cannot fully verify ou

researcharxiv-cs-ai
27 Apr 2026
Model Releases

Towards Temporal Compositional Reasoning in Long-Form Sports Videos

DGX agent

arXiv:2604.22226v1 Announce Type: new Abstract: Sports videos are a challenging domain for multimodal understanding because they involve complex and dynamic human activities. Despite rapid progress in

model-releasesarxiv-cs-cv
27 Apr 2026
Applications

Video Analysis and Generation via a Semantic Progress Function

DGX agent

arXiv:2604.22554v1 Announce Type: new Abstract: Transformations produced by image and video generation models often evolve in a highly non-linear manner: long stretches where the content barely change

applicationsarxiv-cs-cv
27 Apr 2026
Model Releases

I actually switched my personal Claude subscription to this (currently using Mimo v2)

DGX agent

I actually switched my personal Claude subscription to this (currently using Mimo v2) Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from eve

model-releasesnous-research--x
26 Apr 2026
Model Releases

THIS GUY LOST $200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects…

DGX agent

THIS GUY LOST 200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects. it's a system prompt specification file. not some obscure e

model-releasesjeremy-howard--x
26 Apr 2026
Research

2L-LSH: A Locality-Sensitive Hash Function-Based Method For Rapid Point Cloud Indexing

DGX agent

arXiv:2604.21442v1 Announce Type: new Abstract: The development of 3D scanning technology has enabled the acquisition of massive point cloud models with diverse structures and large scales, thereby pr

researcharxiv-cs-cv
24 Apr 2026
Research

A Hybridizable Neural Time Integrator for Stable Autoregressive Forecasting

DGX agent

arXiv:2604.21101v1 Announce Type: new Abstract: For autoregressive modeling of chaotic dynamical systems over long time horizons, the stability of both training and inference is a major challenge in b

researcharxiv-cs-lg
24 Apr 2026
Model Releases

AUDITA: A New Dataset to Audit Humans vs. AI Skill at Audio QA

DGX agent

arXiv:2604.21766v1 Announce Type: new Abstract: Existing audio question answering benchmarks largely emphasize sound event classification or caption-grounded queries, often enabling models to succeed

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Beyond Accuracy: A Stability-Aware Metric for Multi-Horizon Forecasting

DGX agent

arXiv:2601.10863v3 Announce Type: replace Abstract: Traditional time series forecasting methods optimize for accuracy alone. This objective neglects temporal consistency, in other words, how consisten

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

Can MLLMs 'Read' What is Missing?

DGX agent

arXiv:2604.21277v1 Announce Type: new Abstract: We introduce MMTR-Bench, a benchmark designed to evaluate the intrinsic ability of Multimodal Large Language Models (MLLMs) to reconstruct masked text d

model-releasesarxiv-cs-ai
24 Apr 2026
Research

Certified Coil Geometry Learning for Short-Range Magnetic Actuation and Spacecraft Docking Application

DGX agent

arXiv:2507.03806v3 Announce Type: replace-cross Abstract: This paper presents a learning-based framework for approximating an exact magnetic-field interaction model, supported by both numerical and ex

researcharxiv-cs-lg
24 Apr 2026
Safety

Continuous-Utility Direct Preference Optimization

DGX agent

arXiv:2602.00931v2 Announce Type: replace-cross Abstract: Large language model reasoning is often treated as a monolithic capability, relying on binary preference supervision that fails to capture par

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

DAVIS: OOD Detection via Dominant Activations and Variance for Increased Separation

DGX agent

arXiv:2601.22703v2 Announce Type: replace Abstract: Detecting out-of-distribution (OOD) inputs is a critical safeguard for deploying machine learning models in the real world. However, most post-hoc d

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Empirical Comparison of Agent Communication Protocols for Task Orchestration

DGX agent

arXiv:2603.22823v3 Announce Type: replace Abstract: Context. The problem of comparative evaluation of communication protocols for task orchestration by large language model (LLM) agents is considered.

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

EVENT5Ws: A Large Dataset for Open-Domain Event Extraction from Documents

DGX agent

arXiv:2604.21890v1 Announce Type: new Abstract: Event extraction identifies the central aspects of events from text. It supports event understanding and analysis, which is crucial for tasks such as in

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Finding Meaning in Embeddings: Concept Separation Curves

DGX agent

arXiv:2604.21555v1 Announce Type: new Abstract: Sentence embedding techniques aim to encode key concepts of a sentence's meaning in a vector space. However, the majority of evaluation approaches for s

researcharxiv-cs-cl
24 Apr 2026
Research

Frequency-Forcing: From Scaling-as-Time to Soft Frequency Guidance

DGX agent

arXiv:2604.20902v1 Announce Type: cross Abstract: While standard flow-matching models transport noise to data uniformly, incorporating an explicit generation order - specifically, establishing coarse,

researcharxiv-cs-ai
24 Apr 2026
Model Releases

From Codebooks to VLMs: Evaluating Automated Visual Discourse Analysis for Climate Change on Social Media

DGX agent

arXiv:2604.21786v1 Announce Type: new Abstract: Social media platforms have become primary arenas for climate communication, generating millions of images and posts that - if systematically analysed -

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning

DGX agent

arXiv:2604.21495v1 Announce Type: cross Abstract: Numerical reasoning over expert-domain tables often exhibits high in-domain accuracy but limited robustness to domain shift. Models trained with super

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Geometric Characterisation and Structured Trajectory Surrogates for Clinical Dataset Condensation

DGX agent

arXiv:2604.21638v1 Announce Type: new Abstract: Dataset condensation constructs compact synthetic datasets that retain the training utility of large real-world datasets, enabling efficient model devel

model-releasesarxiv-cs-lg
24 Apr 2026
Model Releases

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more aut…

DGX agent

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than any GPT model we've tested, surfacing bugs no

model-releasescognition-ai--x
24 Apr 2026
Applications

Grok Voice is used by @Starlink

DGX agent

Grok Voice is used by @Starlink Introducing Grok Voice Think Fast 1.0 A state-of-the-art voice model built for complex, multi-step workflows with snappy responses and high accuracy. It takes the top s

applicationselon-musk--x
24 Apr 2026
Model Releases

HyperAdapt: Simple High-Rank Adaptation

DGX agent

arXiv:2509.18629v3 Announce Type: replace-cross Abstract: Foundation models excel across diverse tasks, but adapting them to specialized applications often requires fine-tuning, an approach that is me

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Ideological Bias in LLMs' Economic Causal Reasoning

DGX agent

arXiv:2604.21334v1 Announce Type: new Abstract: Do large language models (LLMs) exhibit systematic ideological bias when reasoning about economic causal effects? As LLMs are increasingly used in polic

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Language as a Latent Variable for Reasoning Optimization

DGX agent

arXiv:2604.21593v1 Announce Type: new Abstract: As LLMs reduce English-centric bias, a surprising trend emerges: non-English responses sometimes outperform English on reasoning tasks. We hypothesize t

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Leveraging Multimodal LLMs for Built Environment and Housing Attribute Assessment from Street-View Imagery

DGX agent

arXiv:2604.21102v1 Announce Type: cross Abstract: We present a novel framework for automatically evaluating building conditions nationwide in the United States by leveraging large language models (LLM

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations

DGX agent

arXiv:2509.25868v3 Announce Type: replace Abstract: The mechanisms underlying scientific confabulation in Large Language Models (LLMs) remain poorly understood. We introduce ReFACT (Reddit False And C

model-releasesarxiv-cs-cl
24 Apr 2026
Research

Slot Machines: How LLMs Keep Track of Multiple Entities

DGX agent

arXiv:2604.21139v1 Announce Type: new Abstract: Language models must bind entities to the attributes they possess and maintain several such binding relationships within a context. We study how multipl

researcharxiv-cs-cl
24 Apr 2026
Model Releases

Stealthy Backdoor Attacks against LLMs Based on Natural Style Triggers

DGX agent

arXiv:2604.21700v1 Announce Type: cross Abstract: The growing application of large language models (LLMs) in safety-critical domains has raised urgent concerns about their security. Many recent studie

model-releasesarxiv-cs-ai
24 Apr 2026
Tutorials

Strategic Scaling of Test-Time Compute: A Bandit Learning Approach

DGX agent

arXiv:2506.12721v2 Announce Type: replace Abstract: Scaling test-time compute has emerged as an effective strategy for improving the performance of large language models. However, existing methods typ

tutorialsarxiv-cs-ai
24 Apr 2026
Model Releases

Understanding and Mitigating Spurious Signal Amplification in Test-Time Reinforcement Learning for Math Reasoning

DGX agent

arXiv:2604.21327v1 Announce Type: cross Abstract: Test-time reinforcement learning (TTRL) always adapts models at inference time via pseudo-labeling, leaving it vulnerable to spurious optimization sig

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs

DGX agent

arXiv:2604.21911v1 Announce Type: cross Abstract: Despite impressive progress in capabilities of large vision-language models (LVLMs), these systems remain vulnerable to hallucinations, i.e., outputs

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders

DGX agent

arXiv:2604.19974v1 Announce Type: cross Abstract: Large language models can be uncertain yet correct, or confident yet wrong, raising the question of whether their output-level uncertainty and their a

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

Best Policy Learning from Trajectory Preference Feedback

DGX agent

arXiv:2501.18873v4 Announce Type: replace Abstract: Reinforcement Learning from Human Feedback (RLHF) has emerged as a powerful approach for aligning generative models, but its reliance on learned rew

safetyarxiv-cs-lg
23 Apr 2026
Research

Beyond ZOH: Advanced Discretization Strategies for Vision Mamba

DGX agent

arXiv:2604.20606v1 Announce Type: cross Abstract: Vision Mamba, as a state space model (SSM), employs a zero-order hold (ZOH) discretization, which assumes that input signals remain constant between s

researcharxiv-cs-ai
23 Apr 2026
Model Releases

CRAFT: Training-Free Cascaded Retrieval for Tabular QA

DGX agent

arXiv:2505.14984v2 Announce Type: replace Abstract: Open-Domain Table Question Answering (TQA) involves retrieving relevant tables from a large corpus to answer natural language queries. Traditional d

model-releasesarxiv-cs-cl
23 Apr 2026
Agents

DR-Venus: Towards Frontier Edge-Scale Deep Research Agents with Only 10K Open Data

DGX agent

arXiv:2604.19859v1 Announce Type: cross Abstract: Edge-scale deep research agents based on small language models are attractive for real-world deployment due to their advantages in cost, latency, and

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

Fast-then-Fine: A Two-Stage Framework with Multi-Granular Representation for Cross-Modal Retrieval in Remote Sensing

DGX agent

arXiv:2604.20429v1 Announce Type: new Abstract: Remote sensing (RS) image-text retrieval plays a critical role in understanding massive RS imagery. However, the dense multi-object distribution and com

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

Knowledge Capsules: Structured Nonparametric Memory Units for LLMs

DGX agent

arXiv:2604.20487v1 Announce Type: cross Abstract: Large language models (LLMs) encode knowledge in parametric weights, making it costly to update or extend without retraining. Retrieval-augmented gene

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Mitigating Prompt-Induced Cognitive Biases in General-Purpose AI for Software Engineering

DGX agent

arXiv:2604.16756v2 Announce Type: replace-cross Abstract: Prompt-induced cognitive biases are changes in a general-purpose AI (GPAI) system's decisions caused solely by biased wording in the input (e.

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

DGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

model-releasesallie-k--miller--x
23 Apr 2026
Model Releases

Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL

DGX agent

arXiv:2604.20835v1 Announce Type: new Abstract: Modern language models demonstrate impressive coding capabilities in common programming languages (PLs), such as C++ and Python, but their performance i

model-releasesarxiv-cs-cl
23 Apr 2026
← Previous
1…407408409410411…1326
Next →