AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,574 results
Model Releases

Pretraining Induces a Reusable Spectral Basis for Downstream Task Adaptation

DGX agent

arXiv:2605.07302v1 Announce Type: new Abstract: Finetuning pretrained models occurs in a low-dimensional subspace of the full parameter space. Prior work has focused on characterizing this optimizatio

model-releasesarxiv-cs-lg
11 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

PRIMED: Adaptive Modality Suppression for Referring Audio-Visual Segmentation via Biased Competition

DGX agent

arXiv:2605.07154v1 Announce Type: new Abstract: Referring Audio-Visual Segmentation (Ref-AVS) seeks to localize and segment target objects in video frames based on visual, auditory, and textual referr

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ProcObject-10K: Benchmarking Object-Centric Procedural Understanding in Instructional Videos

DGX agent

arXiv:2512.03479v2 Announce Type: replace Abstract: Procedural activities are fundamentally driven by object state transitions, yet existing instructional video benchmarks remain action-centric and ca

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study

DGX agent

arXiv:2605.07422v1 Announce Type: cross Abstract: Qualitative analysis plays a pivotal role in understanding the human and social aspects of software engineering. However, it remains a demanding proce

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

PSK@EEUCA 2026: Fine-Tuning Large Language Models with Synthetic Data Augmentation for Multi-Class Toxicity Detection in Gaming Chat

DGX agent

arXiv:2605.07201v1 Announce Type: cross Abstract: This paper describes our system for the EEUCA 2026 Shared Task on Understanding Toxic Behavior in Gaming Communities. The task involves classifying Wo

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Quality-Conditioned Agreement in Automated Short Answer Scoring: Mid-Range Degradation and the Impact of Task-Specific Adaptation

DGX agent

arXiv:2605.07647v1 Announce Type: cross Abstract: Automated short answer scoring (ASAS) is shifting from discriminative, fine-tuned models to large language models (LLMs) used in few-shot settings. Th

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Query-efficient model evaluation using cached responses

DGX agent

arXiv:2605.07096v1 Announce Type: cross Abstract: Evaluating a new model on an existing benchmark is often necessary to understand its behavior before deployment. For modern evaluation frameworks, gen

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Quotient Semivalues for False-Name-Resistant Data Attribution

DGX agent

arXiv:2605.07663v1 Announce Type: cross Abstract: Data valuation methods allocate payments and audit training data's contribution to machine-learning pipelines; however, they often assume passive cont

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to …

DGX agent

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens an

model-releasesnous-research--x
11 May 2026
Model Releases

Qwen3-VL-Seg: Unlocking Open-World Referring Segmentation with Vision-Language Grounding

DGX agent

arXiv:2605.07141v1 Announce Type: cross Abstract: Open-world referring segmentation requires grounding unconstrained language expressions to precise pixel-level regions. Existing multimodal large lang

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Randomness is sometimes necessary for coordination

DGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation

DGX agent

arXiv:2605.07334v1 Announce Type: new Abstract: Video Reasoning Segmentation (VRS) aims to segment target objects in videos based on implicit instructions that convey human intent and temporal logic.

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Read the full blog: https://www.together.ai/blog/serving-deepseek-v4-why-million-token-context-is-an-inference-systems-problem# Watch the fu…

DGX agent

Read the full blog: https://www.together.ai/blog/serving-deepseek-v4-why-million-token-context-is-an-inference-systems-problem# Watch the full webinar on DeepSeek v4: https://www.youtube.com/watch?v=D

model-releasestogether-ai--x
11 May 2026
Model Releases

Real-IAD MVN: A Multi-View Normal Vector Dataset and Benchmark for High-Fidelity Industrial Anomaly Detection

DGX agent

arXiv:2605.07149v1 Announce Type: new Abstract: Industrial Anomaly Detection (IAD) is critical for quality control, but existing methods struggle with subtle, geometric defects. Standard 2D (RGB) imag

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

ReasonSTL: Bridging Natural Language and Signal Temporal Logic via Tool-Augmented Process-Rewarded Learning

DGX agent

arXiv:2605.06483v2 Announce Type: replace Abstract: Signal Temporal Logic (STL) is an expressive formal language for specifying spatio-temporal requirements over real-valued, real-time signals. It has

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion

DGX agent

arXiv:2503.06223v5 Announce Type: replace Abstract: Large Vision-Language Models (VLMs) are increasingly deployed in open-ended environments, where ensuring reliable safety under multimodal inputs is

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Region4Web: Rethinking Observation Space Granularity for Web Agents

DGX agent

arXiv:2605.07134v1 Announce Type: cross Abstract: Web agents perceive web pages through an observation space, yet its granularity has remained an underexamined design choice. Existing work treats obse

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Rep2Text: Decoding Full Text from a Single LLM Token Representation

DGX agent

arXiv:2511.06571v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress across diverse tasks, yet their internal mechanisms remain largely opaque. In t

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

DGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Rethinking Dense Optical Flow without Test-Time Scaling

DGX agent

arXiv:2605.08000v1 Announce Type: new Abstract: Recent progress in dense optical flow has been driven by increasingly complex architectures and multi-step refinement for test-time scaling. While these

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Rethinking Weight Tying: Pseudo-Inverse Tying for LM Stable Training and Updates

DGX agent

arXiv:2602.04556v2 Announce Type: replace Abstract: Weight tying is widely used in compact language models to reduce parameters by sharing the token table between the input embedding and the output pr

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Retina-RAG: Retrieval-Augmented Vision-Language Modeling for Joint Retinal Diagnosis and Clinical Report Generation

DGX agent

arXiv:2605.06173v2 Announce Type: replace-cross Abstract: Diabetic Retinopathy (DR) is a leading cause of preventable blindness among working-age adults worldwide, yet most automated screening systems

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Revisiting Transformer Layer Parameterization Through Causal Energy Minimization

DGX agent

arXiv:2605.07588v1 Announce Type: cross Abstract: Transformer blocks typically combine multi-head attention (MHA) for token mixing with gated MLPs for token-wise feature transformation, yet many choic

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Robust Sublinear Convergence Rates for Iterative Bregman Projections

DGX agent

arXiv:2602.01372v2 Announce Type: replace-cross Abstract: Entropic regularization provides a simple way to approximate linear programs whose constraints split into two or more tractable blocks. The re

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning

DGX agent

arXiv:2605.08061v1 Announce Type: new Abstract: We argue that decomposing reward into weighted, verifiable criteria and using an LLM judge to score them provides a partial-credit optimization signal:

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

S2M-Net: Spectral-Spatial Mixing for Medical Image Segmentation with Morphology-Aware Adaptive Loss

DGX agent

arXiv:2601.01285v2 Announce Type: replace Abstract: Medical image segmentation requires balancing local precision for boundary-critical clinical applications, global context for anatomical coherence,

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

DGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Safe, or Simply Incapable? Rethinking Safety Evaluation for Phone-Use Agents

DGX agent

arXiv:2605.07630v1 Announce Type: cross Abstract: When a phone-use agent avoids harm, does that show safety, or simply inability to act? Existing evaluations often cannot tell. A harmful outcome may b

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Safety Anchor: Defending Harmful Fine-tuning via Geometric Bottlenecks

DGX agent

arXiv:2605.05995v2 Announce Type: replace-cross Abstract: The safety alignment of Large Language Models (LLMs) remains vulnerable to Harmful Fine-tuning (HFT). While existing defenses impose constrain

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Sat3R: Satellite DSM Reconstruction via RPC-Aware Depth Fine-tuning

DGX agent

arXiv:2605.07264v1 Announce Type: new Abstract: Accurate Digital Surface Model (DSM) reconstruction from satellite imagery is critical for applications such as disaster response, urban planning, and l

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SatSurfGS: Generalizable 2D Gaussian Splatting for Sparse-View Satellite Surface Reconstruction

DGX agent

arXiv:2605.07181v1 Announce Type: new Abstract: Sparse-view satellite image surface reconstruction remains highly challenging, fundamentally because the reliability of multi-view matching under satell

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Scaling Categorical Flow Maps

DGX agent

arXiv:2605.07820v1 Announce Type: new Abstract: Continuous diffusion and flow matching models could represent a powerful alternative to autoregressive approaches for language modelling (LM), as they u

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts

DGX agent

arXiv:2602.03473v2 Announce Type: replace-cross Abstract: Continual learning, especially class-incremental learning (CIL), on the basis of a pre-trained model (PTM) has garnered substantial research i

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SCENE: Recognizing Social Norms and Sanctioning in Group Chats

DGX agent

arXiv:2605.07823v1 Announce Type: new Abstract: Online group chats are social spaces with implicit behavior patterns that, when broken, are often met with social sanctioning from the group. The abilit

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

DGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ScrapeGraphAI-100k: Dataset for Schema-Constrained LLM Generation

DGX agent

arXiv:2602.15189v2 Announce Type: replace-cross Abstract: Producing output that conforms to a specified JSON schema underlies tool use, structured extraction, and knowledge base construction in modern

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SeedPolicy: Horizon Scaling via Self-Evolving Diffusion Policy for Robot Manipulation

DGX agent

arXiv:2603.05117v3 Announce Type: replace Abstract: Imitation Learning (IL) enables robots to acquire manipulation skills from expert demonstrations. Diffusion Policy (DP) models multi-modal expert be

model-releasesarxiv-cs-ro
11 May 2026
Model Releases

Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback

DGX agent

arXiv:2605.07977v1 Announce Type: new Abstract: Recent works have advanced feedback-based learning systems, whereby a foundation model is able to intake incoming feedback (e.g., a user) to self-improv

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression

DGX agent

arXiv:2502.01941v3 Announce Type: replace-cross Abstract: While Key-Value (KV) cache compression is essential for efficient LLM inference, current evaluations disproportionately focus on sparse retrie

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SEQUOR: A Multi-Turn Benchmark for Realistic Constraint Following

DGX agent

arXiv:2605.06353v2 Announce Type: replace Abstract: In a conversation, a helpful assistant must reliably follow user directives, even as they refine, modify, or contradict earlier requests. Yet most i

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Setting-Matched and Semantics-Scaled Benchmarking of One-Step Generative Models Against Multistep Diffusion and Flow Models

DGX agent

arXiv:2603.14186v4 Announce Type: replace Abstract: State-of-the-art text-to-image models produce high-quality images, but inference remains expensive as generation requires several sequential ODE or

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Several npm packages for the TanStack web development tools were compromised in the Mini Shai-Hulud supply chain attack; Mistral packages were also affected (Socket)

DGX agent

Socket: Several npm packages for the TanStack web development tools were compromised in the Mini Shai-Hulud supply chain attack; Mistral packages were also affected — - Immediate triage: Run shasum -a

model-releasestechmeme
11 May 2026
Model Releases

ShellfishNet: A Domain-Specific Benchmark for Visual Recognition of Marine Molluscs

DGX agent

arXiv:2605.07338v1 Announce Type: new Abstract: The decline of global shellfish biodiversity poses a severe threat to coastal ecosystems. Although artificial intelligence (AI) technologies show potent

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

DGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

SmellBench: Evaluating LLM Agents on Architectural Code Smell Repair

DGX agent

arXiv:2605.07001v1 Announce Type: cross Abstract: Architectural code smells erode software maintainability and are costly to repair manually, yet unlike localized bugs, they require cross-module reaso

model-releasesarxiv-cs-cl
11 May 2026
← Previous
1…343344345346347…471
Next →