AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Revisiting Parameter-Based Knowledge Editing in Large Language Models: Theoretical Limits and Empirical Evidence

DGX agent

arXiv:2606.00570v1 Announce Type: cross Abstract: Parameter-based knowledge editing updates the internal knowledge of large language models (LLMs) via localized weight modifications and has attracted

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Revisiting Ripple Effects in Knowledge Editing through Pressure-Aware Joint Neighborhood Optimization

DGX agent

arXiv:2606.01610v1 Announce Type: new Abstract: Single-edit updates in large language models can trigger ripple effects across local knowledge neighborhoods: desirable propagation to related facts and

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Riemannian Optimization for Hadamard Products of Low-Rank Matrices

DGX agent

arXiv:2606.01216v1 Announce Type: new Abstract: The elementwise Hadamard product of two low-rank matrices provides a parameter-efficient model for data with multiplicative structure, but its modeling

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

RoboBenchMart: Benchmarking Robots in Retail Environment

DGX agent

arXiv:2511.10276v2 Announce Type: replace-cross Abstract: Most existing robotic manipulation benchmarks focus on tabletop or household scenarios. While these setups have driven impressive progress, it

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RoboSemanticBench: Diagnosing Semantic Grounding in Action Prediction for VLA Models

DGX agent

arXiv:2606.02277v1 Announce Type: new Abstract: Vision-language-action (VLA) models are built on the premise that semantic understanding from pretrained language or vision-language backbones should gu

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

RoboStressBench: Benchmarking VLM Robustness to Physical Visual Stress in Embodied Scenes

DGX agent

arXiv:2606.00828v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong visual understanding and are increasingly deployed in embodied AI systems, where reliable perception und

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation

DGX agent

arXiv:2606.01600v1 Announce Type: cross Abstract: Video world models are increasingly used in robotic manipulation, yet existing benchmarks mostly evaluate them under valid, feasible, and safe instruc

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Robust Learning of a Group DRO Neuron

DGX agent

arXiv:2601.18115v2 Announce Type: replace Abstract: We study the problem of learning a single neuron under standard squared loss in the presence of arbitrary label noise and group-level distributional

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

ROGLE: Robust Global-Local Alignment with Automated Region Supervision for Text-Based Person Search

DGX agent

arXiv:2606.01825v1 Announce Type: new Abstract: Text-Based Person Search (TBPS) aims to retrieve pedestrian images using natural language queries. However, existing TBPS models, especially those based

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use

DGX agent

arXiv:2606.00341v1 Announce Type: cross Abstract: As AI agents are increasingly deployed in real personal and corporate settings (email accounts, development workflows, company databases, etc.), safet

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RoleCDE:Benchmarking and Mitigating Role-Alignment Trade-offs in Role-Playing Agents

DGX agent

arXiv:2606.01552v1 Announce Type: new Abstract: Role-playing agents(RPAs) are widely used to steer large language models(LLMs) toward role-consistent behavior, yet existing benchmarks mainly evaluate

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RPCASSM: Robust PCA State Space Model For Infrared Small Target Detection

DGX agent

arXiv:2606.01689v1 Announce Type: cross Abstract: The detection and segmentation of infrared small targets have important application significance in the fields of surveillance and security, maritime

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

RynnVLA-002: A Unified Vision-Language-Action and World Model

DGX agent

arXiv:2511.17502v3 Announce Type: replace Abstract: We introduce RynnVLA-002, a unified Vision-Language-Action (VLA) and world model. The world model leverages action and visual inputs to predict futu

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Ryze: Evidence-Enriched Data Synthesis from Biomedical Papers

DGX agent

arXiv:2606.00902v1 Announce Type: new Abstract: General-purpose VLMs remain unreliable for biomedical research because valid answers in scientific papers depend on evidence split across figures, table

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

S-SPPO: Semantic-Calibrated Self-Play Preference Optimization

DGX agent

arXiv:2606.01561v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) with human preferences is often formulated via Direct Preference Optimization (DPO). However, the standard Bradley

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation

DGX agent

arXiv:2606.01481v1 Announce Type: new Abstract: With the rapid advancements in text-to-image diffusion models, generative video models (T2V models) like Sora can now produce short synthetic videos fro

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

SafeVLA-Bench: A Benchmark for the Success-Safety Gap in Vision-Language-Action Models

DGX agent

arXiv:2606.00773v1 Announce Type: new Abstract: Vision-language-action (VLA) benchmarks measure whether a policy completes a requested manipulation task, but binary success can hide safety-relevant tr

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Saliency-Aware Model Merging

DGX agent

arXiv:2606.00511v1 Announce Type: cross Abstract: Model merging aims to consolidate multiple task-specific models fine-tuned on different datasets into a unified architecture that performs cross-domai

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Same Payload, Different Channel: Measuring Trust Asymmetry in Tool-Using Language Models

DGX agent

arXiv:2606.00566v1 Announce Type: cross Abstract: As language models take on agentic roles that span calling external APIs, reading tool outputs, and acting on instructions embedded in third-party con

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

DGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders

DGX agent

arXiv:2606.00746v1 Announce Type: new Abstract: Vision foundation models are bottlenecked by the quadratic cost of self-attention, which limits usable resolution and increases the cost of large-scale

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Science Earth: Towards A Planet-Scale Operating System for AI-Native Scientific Discovery

DGX agent

arXiv:2606.01316v1 Announce Type: new Abstract: Scientific discovery demands intelligence, perseverance, and serendipity across vast search spaces. Today, top scientific capabilities remain siloed--on

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Score-Control for Hallucination Reduction in Diffusion Models

DGX agent

arXiv:2606.00377v1 Announce Type: new Abstract: Diffusion models have emerged as the backbone of modern generative AI, powering advances in vision, language, audio and other modalities. Despite their

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

SDR: Set-Distance Rewards for Radiology Report Generation

DGX agent

arXiv:2606.00440v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has rapidly advanced reasoning in vision--language models. However, for chest X-ray report generation, th

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents

DGX agent

arXiv:2606.02302v1 Announce Type: cross Abstract: Autonomous LLM agents increasingly operate in stateful environments where they access tools, files, memory, and external services. While such capabili

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation

DGX agent

arXiv:2603.09292v2 Announce Type: replace-cross Abstract: Measurement of task progress through explicit, actionable milestones is critical for robust robotic manipulation. This progress awareness enab

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

DGX agent

arXiv:2503.06520v3 Announce Type: replace Abstract: Traditional methods for reasoning segmentation rely on supervised fine-tuning with categorical labels and simple descriptions, limiting its out-of-d

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Self-Healing Agentic Orchestrators for Reliable Tool-Augmented Large Language Model Systems

DGX agent

arXiv:2606.01416v1 Announce Type: new Abstract: Tool-augmented large language model (LLM) agents rely on orchestration layers that coordinate planning, retrieval, tool invocation, validation, memory,

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Self-Imitated Diffusion Policy for Efficient and Robust Visual Navigation

DGX agent

arXiv:2601.22965v2 Announce Type: replace Abstract: Diffusion policies (DP) have demonstrated significant potential in visual navigation by capturing diverse multi-modal trajectory distributions. Howe

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

Semi-Supervised Hyperbolic Hierarchical Clustering with Set-Level Structural Priors

DGX agent

arXiv:2606.01525v1 Announce Type: new Abstract: Semi-supervised hierarchical clustering aims to learn a tree structure consistent with data patterns and user-provided supervision. Supervision is usual

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

SENSE: Semantic Embedding Navigation with Soft-gated Evaluation for Retrieval-based Speculative Decoding

DGX agent

arXiv:2606.00021v1 Announce Type: cross Abstract: Speculative Decoding (SD) accelerates Large Language Model (LLM) inference by employing a lightweight draft model to propose candidate tokens, which a

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SentGuard: Sentence-Level Streaming Guardrails for Large Language Models

DGX agent

arXiv:2606.02041v1 Announce Type: new Abstract: Large language models increasingly stream long, reasoning-intensive responses in real time, making when to moderate as critical as whether to moderate.

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition

DGX agent

arXiv:2606.00732v1 Announce Type: new Abstract: Learning long-range non-stationary temporal patterns remains a core challenge for modern sequence models, particularly in strict streaming settings. In

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Sharpness-Aware Hybrid Model Learning for Architecture-Agnostic Parameter Estimation

DGX agent

arXiv:2602.06837v2 Announce Type: replace Abstract: Hybrid modeling, the combination of machine learning models and scientific mathematical models, enables flexible and robust data-driven prediction w

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Short-form Text Rewriting with Phi Silica

DGX agent

arXiv:2606.00462v1 Announce Type: cross Abstract: Short-form text rewriting is a constrained variant of paraphrasing in which limited context and high semantic density leave little room for variation.

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning

DGX agent

arXiv:2603.11653v2 Announce Type: replace Abstract: Continual Reinforcement Learning (CRL) for Vision-Language-Action (VLA) models is a promising direction toward self-improving embodied agents that c

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

SindBERT, the Sailor: Charting the Seas of Turkish NLP

DGX agent

arXiv:2510.21364v2 Announce Type: replace Abstract: Transformer models have revolutionized NLP, yet many morphologically rich languages remain underrepresented in large-scale pre-training efforts. Wit

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Single-Channel Tissue Segmentation via Cross-Modal Distillation from Foundation Models

DGX agent

arXiv:2606.00928v1 Announce Type: new Abstract: Multiplexed fluorescence microscopy improves tissue segmentation by providing complementary channels including nuclear (DAPI) and membrane (E-cadherin),

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

DGX agent

arXiv:2606.01311v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on reusable external skills to solve long-horizon interactive tasks. Existing training-free skill

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SkillHarm: Lifecycle-Aware Skill-Based Attacks via Automated Construction

DGX agent

arXiv:2606.02540v1 Announce Type: new Abstract: Agent skills occupy a privileged position in the agent workflow, as agents are expected to implicitly follow and execute them, rendering third-party ski

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

SkillPager: Query-Adaptive Intra-Skill Navigation via Semantic Node Retrieval

DGX agent

arXiv:2606.00822v1 Announce Type: cross Abstract: Skill-based LLM agents increasingly rely on long procedural documents, but full-document prompting wastes tokens and dilutes information critical to e

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SkyShield: Occupancy as a Safety Interface for Low-Altitude UAV Autonomy

DGX agent

arXiv:2606.00747v1 Announce Type: cross Abstract: For low-altitude Unmanned Aerial Vehicle (UAV) autonomy, 3D spatial understanding is not merely a perception objective, but the safety interface betwe

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning

DGX agent

arXiv:2603.08000v2 Announce Type: replace Abstract: Large reasoning models (LRMs) like OpenAI o1 and DeepSeek-R1 achieve high accuracy on complex tasks by adopting long chain-of-thought (CoT) reasonin

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

SMH-Bench: Benchmarking LLM Agents for Environment-Grounded Reasoning and Action in Smart Homes

DGX agent

arXiv:2606.01912v1 Announce Type: new Abstract: Smart homes are evolving toward complex state-dependent living environments, requiring Large Language Models (LLMs) to reason over user intent, preferen

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

DGX agent

arXiv:2606.02380v1 Announce Type: cross Abstract: As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications,

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Sparse FEONet: A Low-Cost, Memory-Efficient Operator Network via Finite-Element Local Sparsity for Parametric PDEs

DGX agent

arXiv:2601.00672v2 Announce Type: replace-cross Abstract: In this paper, we study the finite element operator network (FEONet), an operator-learning method for parametric problems, originally introduc

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Spatially Distributed Task-Oriented Compression for Multi-Emitter Localization and Characterization with Spectral Overlap

DGX agent

arXiv:2606.01446v1 Announce Type: cross Abstract: Radio frequency spectrum awareness requires the ability to detect, localize, and characterize emitters in dense and contested wireless environments. I

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Spectra-Guided Neural Tucker Factorization

DGX agent

arXiv:2606.00584v1 Announce Type: cross Abstract: This paper proposes Spectra-Guided Neural Tucker Factorization (SG-NTF) for High-Dimensional and Incomplete (HDI) tensor completion. Circumventing dis

model-releasesarxiv-cs-lg
2 Jun 2026
← Previous
1…171172173174175…361
Next →