AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Region4Web: Rethinking Observation Space Granularity for Web Agents

DGX agent

arXiv:2605.07134v1 Announce Type: cross Abstract: Web agents perceive web pages through an observation space, yet its granularity has remained an underexamined design choice. Existing work treats obse

model-releasesarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Rep2Text: Decoding Full Text from a Single LLM Token Representation

DGX agent

arXiv:2511.06571v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress across diverse tasks, yet their internal mechanisms remain largely opaque. In t

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

DGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Rethinking Dense Optical Flow without Test-Time Scaling

DGX agent

arXiv:2605.08000v1 Announce Type: new Abstract: Recent progress in dense optical flow has been driven by increasingly complex architectures and multi-step refinement for test-time scaling. While these

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Rethinking Weight Tying: Pseudo-Inverse Tying for LM Stable Training and Updates

DGX agent

arXiv:2602.04556v2 Announce Type: replace Abstract: Weight tying is widely used in compact language models to reduce parameters by sharing the token table between the input embedding and the output pr

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

DGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization

DGX agent

arXiv:2605.07588v1 Announce Type: cross Abstract: Transformer blocks typically combine multi-head attention (MHA) for token mixing with gated MLPs for token-wise feature transformation, yet many choic

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Robust Sublinear Convergence Rates for Iterative Bregman Projections

DGX agent

arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

DGX agent

arXiv:2605.08061v1 Announce Type: new Abstract: We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal:

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

S2M-Net: Spectral-Spatial Mixing for Medical Image Segmentation with Morphology-Aware Adaptive Loss

DGX agent

arXiv:2601.01285v2 Announce Type: replace Abstract: Medical image segmentation requires balancing local precision for boundary-critical clinical applications, global context for anatomical coherence,

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

DGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents

DGX agent

arXiv:2605.07630v1 Announce Type: cross Abstract: When a phone-use agent avoids harm, does that show safety, or simply inability to act? Existing evaluations often cannot tell. A harmful outcome may b

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Safety Anchor: Defending Harmful Fine-tuning via Geometric Bottlenecks

DGX agent

arXiv:2605.05995v2 Announce Type: replace-cross Abstract: The safety alignment of Large Language Models (LLMs) remains vulnerable to Harmful Fine-tuning (HFT). While existing defenses impose constrain

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Sat3R: Satellite DSM Reconstruction via RPC-Aware Depth Fine-tuning

DGX agent

arXiv:2605.07264v1 Announce Type: new Abstract: Accurate Digital Surface Model (DSM) reconstruction from satellite imagery is critical for applications such as disaster response, urban planning, and l

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SatSurfGS: Generalizable 2D Gaussian Splatting for Sparse-View Satellite Surface Reconstruction

DGX agent

arXiv:2605.07181v1 Announce Type: new Abstract: Sparse-view satellite image surface reconstruction remains highly challenging, fundamentally because the reliability of multi-view matching under satell

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Scaling Categorical Flow Maps

DGX agent

arXiv:2605.07820v1 Announce Type: new Abstract: Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling (LM), as they u

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts

DGX agent

arXiv:2602.03473v2 Announce Type: replace-cross Abstract: Continual learning, especially class-incremental learning (CIL), on the basis of a pre-trained model (PTM) has garnered substantial research i

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SCENE: Recognizing Social Norms and Sanctioning in Group Chats

DGX agent

arXiv:2605.07823v1 Announce Type: new Abstract: Online group chats are social spaces with implicit behavior patterns that, when broken, are often met with social sanctioning from the group. The abilit

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

DGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation

DGX agent

arXiv:2602.15189v2 Announce Type: replace-cross Abstract: Producing output that conforms to a specified JSON schema underlies tool use, structured extraction, and knowledge base construction in modern

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SeedPolicy: Horizon Scaling via Self-Evolving Diffusion Policy for Robot Manipulation

DGX agent

arXiv:2603.05117v3 Announce Type: replace Abstract: Imitation Learning (IL) enables robots to acquire manipulation skills from expert demonstrations. Diffusion Policy (DP) models multi-modal expert be

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback

DGX agent

arXiv:2605.07977v1 Announce Type: new Abstract: Recent works have advanced feedback-based learning systems, whereby a foundation model is able to intake incoming feedback (e.g., a user) to self-improv

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression

DGX agent

arXiv:2502.01941v3 Announce Type: replace-cross Abstract: While Key-Value (KV) cache compression is essential for efficient LLM inference, current evaluations disproportionately focus on sparse retrie

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SEQUOR: A Multi-Turn Benchmark for Realistic Constraint Following

DGX agent

arXiv:2605.06353v2 Announce Type: replace Abstract: In a conversation, a helpful assistant must reliably follow user directives, even as they refine, modify, or contradict earlier requests. Yet most i

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Setting-Matched and Semantics-Scaled Benchmarking of One-Step Generative Models Against Multistep Diffusion and Flow Models

DGX agent

arXiv:2603.14186v4 Announce Type: replace Abstract: State-of-the-art text-to-image models produce high-quality images, but inference remains expensive as generation requires several sequential ODE or

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs

DGX agent

arXiv:2605.07338v1 Announce Type: new Abstract: The decline of global shellfish biodiversity poses a severe threat to coastal ecosystems. Although artificial intelligence (AI) technologies show potent

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

DGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair

DGX agent

arXiv:2605.07001v1 Announce Type: cross Abstract: Architectural code smells erode software maintainability and are costly to repair manually, yet unlike localized bugs, they require cross-module reaso

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Sparse Random-Feature Neural Networks with Krylov-Based SVD for Singularly Perturbed ODE

DGX agent

arXiv:2605.07286v1 Announce Type: cross Abstract: Random-feature neural networks (RFNNs), including architectures with fixed hidden layers and analytically determined output weights, offer fast traini

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

SR^2-LoRA: Self-Rectifying Inter-layer Relations in Low-Rank Adaptation for Class-Incremental Learning

DGX agent

arXiv:2605.07420v1 Announce Type: cross Abstract: Pre-trained models with parameter-efficient fine-tuning (PEFT) have demonstrated promising potential for class-incremental learning (CIL), yet catastr

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios

DGX agent

arXiv:2605.07161v1 Announce Type: new Abstract: AI agents are increasingly used to diagnose and mitigate failures in production systems, known as agentic Site Reliability Engineering (SRE). Current SR

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Star Elastic: Many-in-One Reasoning LLMs with Efficient Budget Control

DGX agent

arXiv:2605.07182v1 Announce Type: new Abstract: Training a family of large language models (LLMs), either from scratch or via iterative compression, is prohibitively expensive and inefficient, requiri

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

SteelDefectX: A Multi-Form Vision-Language Dataset and Benchmark for Steel Surface Defect Analysis

DGX agent

arXiv:2603.21824v2 Announce Type: replace-cross Abstract: Steel surface defect analysis is critical for industrial quality control, yet existing benchmarks rely primarily on label-only annotations, li

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Structure Over Scale: Learning Visual Reasoning from Pedagogical Video

DGX agent

arXiv:2601.23251v2 Announce Type: replace Abstract: State-of-the-art vision-language models (VLMs) score impressively on video benchmarks yet stumble on basic visual reasoning tasks involving spatial

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Structured Prototype-Guided Adaptation for EEG Foundation Models

DGX agent

arXiv:2602.17251v2 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (EFMs) have shown strong potential for transferable representation learning, yet their adaptation in

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Sword: Style-Robust World Models as Simulators via Dynamic Latent Bootstrapping for VLA Policy Post-Training

DGX agent

arXiv:2605.07288v1 Announce Type: cross Abstract: The integration of Vision-Language-Action (VLA) models with World Models has gained increasing attention. One representative approach treats learned W

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SynthForensics: Benchmarking and Evaluating People-Centric Synthetic Video Deepfakes

DGX agent

arXiv:2602.04939v2 Announce Type: replace Abstract: Modern T2V/I2V generators synthesize people increasingly hard to distinguish from authentic footage, while current evaluation suites lag: legacy ben

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

TAG-K: Tail-Averaged Greedy Kaczmarz for Computationally Efficient and Performant Online Inertial Parameter Estimation

DGX agent

arXiv:2510.04839v2 Announce Type: replace Abstract: Accurate online inertial parameter estimation is essential for adaptive robotic control, enabling real-time adjustment to payload changes, environme

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

TajPersLexon: A Tajik-Persian Lexical Resource and Hybrid Model for Cross-Script Low-Resource NLP

DGX agent

arXiv:2605.06886v1 Announce Type: new Abstract: This work introduces TajPersLexon, a curated Tajik--Persian parallel lexical resource of 40,112 word and short-phrase pairs for cross-script lexical ret

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Target-Aware Data Augmentation for SAT Prediction

DGX agent

arXiv:2605.06931v1 Announce Type: new Abstract: Learning-based approaches to NP-hard problems have shown increasing promise, but their progress is fundamentally constrained by the high cost of generat

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

TAS-LoRA: Transformer Architecture Search with Mixture-of-LoRA Experts

DGX agent

arXiv:2605.07256v1 Announce Type: new Abstract: Transformer architecture search (TAS) discovers optimal vision transformer (ViT) architectures automatically, reducing human effort to manually design V

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

TAVIS: A Benchmark for Egocentric Active Vision and Anticipatory Gaze in Imitation Learning

DGX agent

arXiv:2605.07943v1 Announce Type: cross Abstract: Active vision -- where a policy controls its own gaze during manipulation -- has emerged as a key capability for imitation learning, with multiple ind

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent

DGX agent

arXiv:2601.18700v2 Announce Type: replace Abstract: Emotional Support Conversation requires not only affective expression but also grounded instrumental support to provide trustworthy guidance. Howeve

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Teaching Language Models to Think in Code

DGX agent

arXiv:2605.07237v1 Announce Type: new Abstract: Tool-integrated reasoning (TIR) has emerged as a dominant paradigm for mathematical problem solving in language models, combining natural language (NL)

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

TeamBench: Evaluating Agent Coordination under Enforced Role Separation

DGX agent

arXiv:2605.07073v1 Announce Type: new Abstract: Agent systems often decompose a task across multiple roles, but these roles are typically specified by prompts rather than enforced by access controls.

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Test-Time Compute Games

DGX agent

arXiv:2601.21839v2 Announce Type: replace-cross Abstract: Test-time compute has emerged as a promising strategy to enhance the reasoning abilities of large language models (LLMs). However, this strate

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…267268269270271…361
Next →