AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

SigmaScale: LLM Compression with SVD-based Low-Rank Decomposition and Learned Scaling Matrices

DGX agent

arXiv:2606.07098v1 Announce Type: new Abstract: We present SigmaScale, a method for learning auxiliary scaling matrices S to aid truncated Singular Value Decomposition (SVD) based Large Language Model

model-releasesarxiv-cs-cl
8 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Skill-3D: Evolving Scene-Aware Skills for Agentic 3D Spatial Reasoning

DGX agent

arXiv:2606.07436v1 Announce Type: new Abstract: This paper explores agentic 3D spatial understanding, i.e., MLLM agents performing 3D reasoning through tool use. Existing methods often misuse tools an

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Sparse Subspace-to-Expert Sharing for Task-Agnostic Continual Learning

DGX agent

arXiv:2606.07500v1 Announce Type: cross Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to ca

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Spatial-Temporal Decoupled Adapter for Micro-gesture Online Recognition

DGX agent

arXiv:2606.07355v1 Announce Type: new Abstract: Micro-gesture online recognition aims to temporally localize and classify subtle gestures in untrimmed videos. Owing to their extremely short duration,

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Spline Policy: A Structured Representation for Robot Policies

DGX agent

arXiv:2606.07386v1 Announce Type: new Abstract: Modern imitation-learning policies for robot manipulation often represent actions as fixed-resolution action chunks, which are simple and effective but

model-releasesarxiv-cs-ro
8 Jun 2026
Model Releases

STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

DGX agent

arXiv:2606.07036v1 Announce Type: cross Abstract: Synthetic histopathology image generation addresses critical challenges in computational pathology, including patient privacy and the growing need for

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors

DGX agent

arXiv:2606.06891v1 Announce Type: new Abstract: Despite advances in 3D scene understanding, existing 3D Large Multimodal Models operate in offline settings, requiring complete scene observations or pr

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Style or Content? Evaluating Style Classifiers with Controlled Content Overlap

DGX agent

arXiv:2606.07103v1 Announce Type: new Abstract: Style classifiers can use content cues that correlate with style labels in naturally collected data, yet we lack a systematic way to measure this relian

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval

DGX agent

arXiv:2605.06647v2 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue explor

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Classification

DGX agent

arXiv:2606.07479v1 Announce Type: cross Abstract: Turkish idiomatic light verb constructions (LVCs) are challenging for multiword expression processing because they often share the same surface form a

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

SVHighlights: Towards Extremely Long Sport Video Highlight Detection

DGX agent

arXiv:2606.06926v1 Announce Type: new Abstract: While highlight detection for long-form videos is of great practical importance, most existing methods remain limited to short-form content, largely due

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

SW-A^2-Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web

DGX agent

arXiv:2604.04226v2 Announce Type: replace-cross Abstract: The Agentic Web is emerging as a paradigm in which autonomous software agents interact with online resources and with each other to accomplish

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

DGX agent

arXiv:2606.07297v1 Announce Type: cross Abstract: Repository-level coding benchmarks such as SWE-bench have driven a rapid surge in the capabilities of coding agents. Yet they usually treat coding tas

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

TALAN: Task-Aligned Latent Adaptation Networks for Targeted Post-Training of Large Language Models

DGX agent

arXiv:2606.06902v1 Announce Type: new Abstract: Targeted post-training aims to improve reasoning, math, and code without degrading strengths. Low-rank adapters are efficient but task-global; activatio

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Test-Time Trajectory Optimization for Autonomous Driving

DGX agent

arXiv:2606.07170v1 Announce Type: new Abstract: End-to-end planners for autonomous driving typically generate a set of candidate trajectories, score each one, and return the highest-scoring candidate.

model-releasesarxiv-cs-ro
8 Jun 2026
Model Releases

TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment

DGX agent

arXiv:2606.07451v1 Announce Type: cross Abstract: Vision-language models such as CLIP are highly useful for diverse tasks due to their shared image-text embedding space. Despite this, the image and te

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Textual Supervision Enhances Geospatial Representations in Vision-Language Models

DGX agent

arXiv:2606.07172v1 Announce Type: cross Abstract: Geospatial understanding is a critical yet underexplored dimension in the development of machine learning systems for tasks such as image geolocation

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

The Fine-Tuning Trap: Evaluating Negative Transfer and the Role of PEFT in Sub-1B Mathematical Reasoning

DGX agent

arXiv:2606.06920v1 Announce Type: cross Abstract: Deploying Small Language Models (SLMs) on edge devices requires efficient fine-tuning strategies that adapt models to new tasks without degrading thei

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

The Geometry of Representational Failures in Vision Language Models

DGX agent

arXiv:2602.07025v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) exhibit puzzling failures in multi-object visual tasks, such as hallucinating non-existent elements or failing t

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

The Piggyback Hypothesis of Generalization: Explaining and Mitigating Emergent Misalignment

DGX agent

arXiv:2606.06667v1 Announce Type: new Abstract: The mechanisms behind LLMs' broad over-generalization beyond training examples remain unclear. Emergent misalignment (EM) offers a striking case study:

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

The Post-GCN Decade Revisited: Curvature-Stratified Evaluation of Relational Learning

DGX agent

arXiv:2606.06397v2 Announce Type: replace Abstract: Current evaluation practices in relational learning rely heavily on flat leaderboards that average performance across heterogeneous datasets, implic

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

DGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Think Like a Pilot: Fine-Grained Long-Horizon UAV Navigation

DGX agent

arXiv:2606.06836v1 Announce Type: cross Abstract: Language-guided UAV agents must execute long-horizon semantic instructions while producing smooth, physically feasible continuous flight commands, yet

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning

DGX agent

arXiv:2606.06915v1 Announce Type: cross Abstract: Test-time compute (TTC) scaling has emerged as a powerful paradigm for improving large language model (LLM) reasoning by allocating additional compute

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

TokaMind: A Multi-Modal Transformer Foundation Model for Tokamak Plasma Dynamics

DGX agent

arXiv:2602.15084v2 Announce Type: replace-cross Abstract: We present TokaMind, to our knowledge the first open-source foundation model for tokamak plasma dynamics, based on a Multi-Modal Transformer (

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Towards Tight Bounds for Streaming Attention

DGX agent

arXiv:2606.07205v1 Announce Type: cross Abstract: The attention mechanism is a cornerstone of modern transformer architectures. However, its expressive power comes at the cost of quadratic runtime and

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Trading Engagement for Sustainability: Carbon-Aware Re-ranking for E-commerce Recommendations

DGX agent

arXiv:2606.04550v1 Announce Type: cross Abstract: E-commerce recommender systems strongly influence which products users consider and purchase, yet sustainability signals such as Product Carbon Footpr

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Tree-of-Experience: A Structured Experience-Management Solution for Self-Evolving Agents under Low-Repetition and Implicit-Reward Environments

DGX agent

arXiv:2606.06960v1 Announce Type: new Abstract: Experience-based self-evolution is crucial for LLM agents, but existing benchmarks often assume explicit goals, stable task patterns, and clear feedback

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

TSAQA: Time Series Analysis Question And Answering Benchmark

DGX agent

arXiv:2601.23204v2 Announce Type: replace Abstract: Time series data are integral to critical applications across domains such as finance, healthcare, transportation, and environmental science. While

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Twin: Tuning Learning Rate and Weight Decay of Deep Homogeneous Classifiers without Validation

DGX agent

arXiv:2403.05532v2 Announce Type: replace-cross Abstract: We introduce Tune without Validation (Twin), a simple and effective pipeline for tuning learning rate and weight decay of homogeneous classifi

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Uncertainty-Aware LLM-Guided Policy Shaping for Sparse-Reward Reinforcement Learning

DGX agent

arXiv:2606.06673v1 Announce Type: new Abstract: Sparse rewards and heterogeneous task sequences remain persistent challenges in Reinforcement Learning (RL), often resulting in slow convergence, weak g

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Uncertainty-Guided Label Rebalancing for CPS Safety Monitoring

DGX agent

arXiv:2603.25670v3 Announce Type: replace Abstract: Safety monitoring is essential for Cyber-Physical Systems (CPSs). However, unsafe events are rare in real-world CPS operations, creating an extreme

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

Uniform Stability and Generalization Error of GD and SGD on Fixed-Point Parameters

DGX agent

arXiv:2606.06934v1 Announce Type: new Abstract: We analyze generalization error, uniform stability, and uniform argument stability of gradient descent (GD) and stochastic gradient descent (SGD) over d

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

UniSHARP: Universal Sharp Monocular View Synthesis

DGX agent

arXiv:2606.07514v1 Announce Type: new Abstract: In this work, we focus on extending SHARP, the popular photorealistic view synthesis method, for universal monocular rendering across a continuum of cam

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

UnpredictaBench: A Benchmark for Evaluating Distributional Randomness in LLMs

DGX agent

arXiv:2606.06622v1 Announce Type: new Abstract: We introduce UnpredictaBench, an evaluation that tests the ability of large language models (LLMs) to capture true underlying distributions. As LLMs are

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Unsupervised Continual Clustering via Forward-Backward Knowledge Distillation

DGX agent

arXiv:2606.07474v1 Announce Type: new Abstract: Unsupervised Continual Learning (UCL) aims to enable neural networks to learn sequential tasks without labels or access to past data. A major challenge

model-releasesarxiv-cs-lg
8 Jun 2026
Model Releases

UrduMMLU: A Massive Multitask Benchmark for Urdu Language Understanding

DGX agent

arXiv:2606.07167v1 Announce Type: cross Abstract: Meaningful multilingual evaluation must test models in the target language and educational context. Urdu, spoken by more than 230 million people, lack

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

VeriDrive: Verifiable Counterfactual Supervision for Cost-Efficient Vision-Language Planning

DGX agent

arXiv:2606.07338v1 Announce Type: new Abstract: Vision-language driving models increasingly use reasoning supervision to bridge perception, prediction, and planning, but existing driving rationales ar

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

VideoSEG-O3: A Multi-turn Reinforcement Learning Framework for Reasoning Video Object Segmentation

DGX agent

arXiv:2606.06819v1 Announce Type: new Abstract: Reasoning Video Object Segmentation (RVOS) demands a sophisticated integration of temporal dynamics, spatial details, and linguistic reasoning to achiev

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

What Your Posts Reveal: A Benchmark and Agentic Framework for User-Level Privacy Leakage on Social Media

DGX agent

arXiv:2606.06784v1 Announce Type: cross Abstract: Public social media posts can reveal private information through weak cues scattered across text, images, or metadata. Such leakage is often cumulativ

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations

DGX agent

arXiv:2606.07237v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in healthcare for tasks such as clinical question answering, diagnosis support, and report summariz

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

When Recovery Matters: The Blind Spot of Surrogate Privacy in MLLM Editing

DGX agent

arXiv:2606.07171v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) enable flexible instruction-driven image editing, but privacy risks arise when user images expose diverse and u

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

When to Think Deeply: Inhibitory Deliberation for LLM Reasoning

DGX agent

arXiv:2606.06745v1 Announce Type: new Abstract: Reasoning Large Language Models can improve problem-solving performance through deliberative inference, but invoking slow reasoning for every input is c

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

Which Anatomy Matters Under Limited Labels? A Data-Efficient Anatomy-Aware Benchmark for Cardiac Pathology Prediction

DGX agent

arXiv:2606.06509v1 Announce Type: cross Abstract: Numerous medical imaging problems must be solved under limited labels and constrained compute, yet it remains unclear whether performance gains are dr

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

WorldBench: A Challenging and Visually Diverse Multimodal Reasoning Benchmark

DGX agent

arXiv:2606.06538v1 Announce Type: new Abstract: In real-world applications, models are expected to perform reliably across diverse settings. Yet, many existing multimodal benchmarks expand task types

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs

DGX agent

arXiv:2601.12359v1 Announce Type: cross Abstract: Prompt injection attacks have become an increasing vulnerability for LLM applications, where adversarial prompts exploit indirect input channels such

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

ADK Arena: Evaluating Agent Development Kits via LLM-as-a-Developer

DGX agent

arXiv:2606.05548v1 Announce Type: cross Abstract: The rapid proliferation of Agent Development Kits (ADKs), SDK-level frameworks for building LLM-powered autonomous agents, has outpaced any empirical

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads

DGX agent

arXiv:2606.06448v1 Announce Type: new Abstract: LLM agents are increasingly deployed on long-horizon tasks requiring sustained reasoning over extended interaction histories. Realizing this at scale re

model-releasesarxiv-cs-ai
6 Jun 2026
← Previous
1…149150151152153…361
Next →