AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

LifeSide: Benchmarking Agents as Lifelong Digital Companions

DGX agent

arXiv:2606.04660v1 Announce Type: new Abstract: Lifelong digital companions must integrate cross-session cues, continually update their understanding of users, and adapt to shifting privacy boundaries

model-releasesarxiv-cs-cl
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection

DGX agent

arXiv:2606.04050v1 Announce Type: cross Abstract: Existing quantization methods are fundamentally limited by rigid, integer-based bit-widths (e.g., 2, 3-bit), resulting in a ``deployment gap' where La

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

LimiX-2M: Mitigating Low-Rank Collapse and Attention Bottlenecks in Tabular Foundation Models

DGX agent

arXiv:2606.04485v1 Announce Type: new Abstract: Tabular foundation models (TFMs) increasingly rival tree ensembles, but their performance is often compute-inefficient: with standard affine scalar toke

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Literature-Guided Minimax Optimization of Virtual Epilepsy Neurostimulation

DGX agent

arXiv:2606.04339v1 Announce Type: new Abstract: Computational models of epilepsy promise patient-specific treatment design, but most optimization workflows still search for parameters that perform wel

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

LLMs + Persona-Plug = Personalized LLMs

DGX agent

arXiv:2409.11901v2 Announce Type: replace Abstract: Personalization plays a critical role in numerous language tasks and applications, since users with the same requirements may prefer diverse outputs

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit

DGX agent

arXiv:2606.04274v1 Announce Type: new Abstract: As large language models (LLMs) become default tools for online information verification, an implicit assumption follows them: that scale and general ca

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Longer Context, Deeper Thinking: Uncovering the Role of Long-Context Ability in Reasoning

DGX agent

arXiv:2505.17315v2 Announce Type: replace Abstract: Recent language models exhibit strong reasoning capabilities, yet the influence of long-context capacity on reasoning remains underexplored. In this

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling

DGX agent

arXiv:2606.04438v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, main

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

M^3Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks

DGX agent

arXiv:2606.05008v1 Announce Type: cross Abstract: As multi-modal models advance towards long-form video understanding, memory emerges as a critical capability. Despite substantial efforts in developin

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MedForge: Interpretable Medical Deepfake Detection via Forgery-aware Reasoning

DGX agent

arXiv:2603.18577v2 Announce Type: replace Abstract: Text-guided image editors can now manipulate authentic medical scans with high fidelity, enabling lesion implantation/removal that threatens clinica

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MemoryDocDataSet: A Benchmark for Joint Conversational Memory and Long Document Reasoning

DGX agent

arXiv:2606.04442v1 Announce Type: cross Abstract: AI systems increasingly need to combine two demanding capabilities: navigating multi-session conversation history and performing deep reading comprehe

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MesaNet: Sequence Modeling by Locally Optimal Test-Time Training

DGX agent

arXiv:2506.05233v2 Announce Type: replace-cross Abstract: Sequence modeling is currently dominated by causal transformer architectures that use softmax self-attention. Although widely adopted, transfo

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MeshTok: Efficient Multi-Scale Tokenization for Scalable PDE Transformers

DGX agent

arXiv:2606.04366v1 Announce Type: new Abstract: Conventional patchified Transformers operate on uniform spatial partitions, distributing computational effort evenly across the domain irrespective of l

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Metric-Aware Hybrid Forecasting for the CTF4Science Lorenz Challenge

DGX agent

arXiv:2606.04191v1 Announce Type: cross Abstract: We describe our approach to the CTF4Science Lorenz challenge, a benchmark that mixes short-horizon forecasting, long-time distribution matching, and t

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments

DGX agent

arXiv:2606.04171v1 Announce Type: cross Abstract: File-type classification underlies many workflows like malware triage, forensic carving, packet inspection, and storage indexing. Learned systems such

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

MineXplore: An Open-Source Reinforcement Learning Exploration Benchmark for GNSS-Denied Underground Environment

DGX agent

arXiv:2606.04569v1 Announce Type: new Abstract: Underground mines present extreme conditions for autonomous robot navigation: GPS is denied, lighting is degraded, and tunnel topology is loop-rich and

model-releasesarxiv-cs-ro
4 Jun 2026
Model Releases

Multi-Column RBF Neural Network Using Adaptive and Non-Adaptive Particle Swarm Optimization

DGX agent

arXiv:2606.05150v1 Announce Type: cross Abstract: The radial basis function neural network (RBFN) trained with a gradient descending algorithm provides an effective fully connected structure in both s

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge

DGX agent

arXiv:2606.04581v1 Announce Type: cross Abstract: Speculative inference (SPIN) was originally developed as an efficient architecture to accelerate Large Language Models (LLMs). In this work, we propos

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Need to Know: Contextual-Integrity-Grounded Query Rewriting for Privacy-Conscious LLM Delegation

DGX agent

arXiv:2606.04067v1 Announce Type: cross Abstract: As LLMs become increasingly woven into everyday workflows, user queries sent to cloud hosted LLMs routinely mix task-essential content with task non-e

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Neural Galerkin Normalizing Flows for Bayesian Inference of Diffusions with Inaccessible Boundaries

DGX agent

arXiv:2606.04324v1 Announce Type: new Abstract: One of the primary challenges in Bayesian inference on the parameters of a diffusion model from discrete observations is the unavailability of an analyt

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

DGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

NextMotionQA: Benchmarking and Judging Human Motion Understanding with Vision-Language Models

DGX agent

arXiv:2606.04773v1 Announce Type: cross Abstract: Reliable evaluation of human motion understanding is fundamental to advancing embodied AI, robotics, and animation. However, existing benchmarks suffe

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

NoRA: Evaluating Grounded Reasonableness in Visual First-person Normative Action Reasoning

DGX agent

arXiv:2606.04806v1 Announce Type: cross Abstract: LLMs and agentic systems are increasingly deployed in social environments, making normative competence critical for safe and appropriate behavior. How

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

DGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

OckBench: Measuring the Efficiency of LLM Reasoning

DGX agent

arXiv:2511.05722v3 Announce Type: replace-cross Abstract: Large language models (LLMs) such as GPT-5 and Gemini 3 have pushed the frontier of automated reasoning and code generation. Yet current bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

On the Relationship Between CoCoA and ADMM for Distributed Empirical Risk Minimization

DGX agent

arXiv:2502.00470v3 Announce Type: replace-cross Abstract: Distributed empirical risk minimization (ERM) is often studied through two influential yet seemingly separate families of methods: CoCoA-type

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

DGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Optical-Guided Neural Collapse for SAR Few-Shot Class Incremental Learning

DGX agent

arXiv:2606.04528v1 Announce Type: cross Abstract: Few-shot class-incremental learning (FSCIL) in synthetic aperture radar imagery presents unique challenges due to severe data scarcity and SAR-specifi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Parameter-Efficient Fine-Tuning with Learnable Rank

DGX agent

arXiv:2606.04325v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) is a popular parameter-efficient fine-tuning (PEFT) method that restricts weight updates to low-rank adapters, introducing a

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

PE-MHL: Physics-Encoded Modular Hybrid Layers for Scalable Learning of Complex Systems

DGX agent

arXiv:2606.04290v1 Announce Type: new Abstract: Hybrid models that combine physics-based and data-driven components have shown strong potential for achieving accuracy and interpretability in control a

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

DGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Physics-Informed Video Generation via Mixture-of-Experts Latent Alignment

DGX agent

arXiv:2606.04737v1 Announce Type: new Abstract: Large-scale video generation models have made remarkable progress in semantic consistency and visual quality, producing videos that are increasingly coh

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Plan, Watch, Recover: A Benchmark and Architectures for Proactive Procedural Assistance

DGX agent

arXiv:2606.04970v1 Announce Type: cross Abstract: We envision a proactive multi-modal assistant system which gives users real-time step-by-step guidance on a procedural task, autonomously deciding ext

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

DGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Prompt-Level Distillation: A Non-Parametric Alternative to Model Fine-Tuning for Efficient Reasoning

DGX agent

arXiv:2602.21103v2 Announce Type: replace Abstract: Advanced reasoning typically requires Chain-of-Thought prompting, which is accurate but incurs prohibitive latency and substantial test-time inferen

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

DGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Provably Reduced Sample Cost in Prior-Guided Hyperparameter Optimization

DGX agent

arXiv:2606.04866v1 Announce Type: new Abstract: Large-scale hyperparameter optimization (HPO) in automated machine learning (AutoML) consumes substantial computational resources, raising growing conce

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Pseudospectral Bounds for Transient Amplification in Coupled Gradient Descent

DGX agent

arXiv:2606.04031v1 Announce Type: new Abstract: Coupled gradient descent--where the update of one parameter block depends on another--underlies bilevel optimization, two-time-scale stochastic approxim

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

QO-Bench: Diagnosing Query-Operator-Preserving Retrieval over Typed Event Tuples

DGX agent

arXiv:2606.04646v1 Announce Type: cross Abstract: Many real-world questions over business, legal, and scientific corpora are natural-language versions of database-style queries over records latent in

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

QPredSGG: Hybrid Quantum Predicate Learning for Long-Tailed Scene Graph Generation

DGX agent

arXiv:2606.04689v1 Announce Type: cross Abstract: Scene Graph Generation (SGG) requires relational reasoning over objects and their interactions, but performance is often limited by severe long-tail p

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Quantum entanglement provides a competitive advantage in adversarial games

DGX agent

arXiv:2603.10289v2 Announce Type: replace-cross Abstract: Whether uniquely quantum resources confer advantages in fully classical, competitive environments remains an open question. Competitive zero-s

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

QuBLAST: A Framework for Quantizing Large Language Models with Block-Level Compression Approach and Activation Scaling Strategy

DGX agent

arXiv:2606.04620v1 Announce Type: cross Abstract: LLMs have become the state-of-the-art algorithms for solving NLP tasks. However, they typically come at huge computational and memory costs, thus maki

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

RAMPART: Registry-based Agentic Memory with Priority-Aware Runtime Transformation

DGX agent

arXiv:2606.04628v1 Announce Type: new Abstract: RAMPART is a compile-time memory model and pure in-RAM block registry for LLM-based agents. Context assembly is a programmable runtime operation where c

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression

DGX agent

arXiv:2511.21035v2 Announce Type: replace Abstract: Holography offers significant potential for AR/VR applications. However, its adoption is limited by the high demand for data compression. Existing d

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Reasoning over Boundaries: Enhancing Specification Alignment via Test-time Deliberation

DGX agent

arXiv:2509.14760v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in diverse real-world scenarios, each governed by bespoke behavioral and safety specifications

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Revisiting Vul-RAG: Reproducibility and Replicability of RAG-based Vulnerability Detection with Open-Weight Models

DGX agent

arXiv:2606.04739v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong potential for automated software vulnerability detection, particularly in retrieval-augmented generatio

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

RIDE: An Open Dataset and Benchmark for Train Delay Prediction

DGX agent

arXiv:2606.05070v1 Announce Type: new Abstract: Train delay prediction is an important problem for both passengers and railway operators, yet progress in the field remains difficult to assess due to t

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Robust Multi-view Clustering against Imperfect Information

DGX agent

arXiv:2606.04343v1 Announce Type: new Abstract: Real-world multi-view data always suffer from imperfect information problem, where the view-specific observations are absent (i.e., Incomplete Views, IV

model-releasesarxiv-cs-cv
4 Jun 2026
← Previous
1…156157158159160…361
Next →