AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

RISK: A Framework for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2509.21982v2 Announce Type: replace Abstract: E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scrap

model-releasesarxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RiTeK: A Dataset for Large Language Models Complex Reasoning over Textual Knowledge Graphs in Medicine

DGX agent

arXiv:2410.13987v3 Announce Type: replace Abstract: Answering complex real-world questions in the medical domain often requires accurate retrieval from medical Textual Knowledge Graphs (medical TKGs),

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

RL-Driven Sustainable Land-Use Allocation for the Lake Malawi Basin

DGX agent

arXiv:2604.03768v2 Announce Type: replace Abstract: Unsustainable land-use practices in ecologically sensitive regions threaten biodiversity, water resources, and the livelihoods of millions. This pap

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RL makes MLLMs see better than SFT

DGX agent

arXiv:2510.16333v2 Announce Type: replace Abstract: A dominant assumption in Multimodal Language Model (MLLM) research is that its performance is largely inherited from the LLM backbone, given its imm

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

DGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Robust Adversarial Policy Optimization Under Dynamics Uncertainty

DGX agent

arXiv:2604.10974v1 Announce Type: new Abstract: Reinforcement learning (RL) policies often fail under dynamics that differ from training, a gap not fully addressed by domain randomization or existing

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Robust Fair Disease Diagnosis in CT Images

DGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RobustMedSAM: Degradation-Resilient Medical Image Segmentation via Robust Foundation Model Adaptation

DGX agent

arXiv:2604.09814v1 Announce Type: new Abstract: Medical image segmentation models built on Segment Anything Model (SAM) achieve strong performance on clean benchmarks, yet their reliability often degr

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo

DGX agent

arXiv:2505.09368v2 Announce Type: replace Abstract: Standard benchmarks for optical flow, scene flow, and stereo vision algorithms generally focus on model accuracy rather than robustness to image cor

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight

DGX agent

arXiv:2512.19691v3 Announce Type: replace Abstract: Reference labels for machine-learning benchmarks are increasingly synthesized with LLM assistance, but their reliability remains underexamined. We a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Scaling unstructured enterprise knowledge with BigQuery Graph, and Kineviz GraphXR

DGX agent

Over 80% of enterprise data lives in unstructured form — PDFs, emails, reports, regulatory filings. Most of the time, such sources contain critical business information, yet they remain difficult to a

model-releasesgoogle-cloud-ai
14 Apr 2026
Model Releases

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

DGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions

DGX agent

arXiv:2307.01139v2 Announce Type: replace-cross Abstract: Instruction finetuning is a popular paradigm to align large language models (LLM) with human intent. Despite its popularity, this idea is less

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

DGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

DGX agent

arXiv:2512.12675v2 Announce Type: replace-cross Abstract: Subject-driven image generation has advanced from single- to multi-subject composition, while neglecting distinction, the ability to distingui

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding

DGX agent

arXiv:2604.11244v1 Announce Type: new Abstract: Advances in Multimodal Large Language Models (MLLMs) are transforming video captioning from a descriptive endpoint into a semantic interface for both vi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Sectigo launches Private PQC to enable post-quantum certificate testing in existing workflows

DGX agent

Digital certificates and certificate lifecycle management firm Sectigo Ltd. today announced the launch of Private PQC, a new feature that allows enterprises to issue and manage private post-quantum cr

model-releasessiliconangle
14 Apr 2026
Model Releases

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

DGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Securing the AI era across the public sector

DGX agent

The agentic era demands a new security era. Our mission is to be the world’s most trusted security partner in this new era, helping every organization accelerate their security transformation with the

model-releasesgoogle-cloud-ai
14 Apr 2026
Model Releases

Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking

DGX agent

arXiv:2604.10299v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely on attention-based retrieval of safety instructions to maintain alignment during generation. Existing attack

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models

DGX agent

arXiv:2505.15489v4 Announce Type: replace-cross Abstract: The impact of multimodal misinformation arises not only from factual inaccuracies but also from the misleading narratives that creators delibe

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Seeing Through the Tool: A Controlled Benchmark for Occlusion Robustness in Foundation Segmentation Models

DGX agent

arXiv:2604.11711v1 Announce Type: new Abstract: Occlusion, where target structures are partially hidden by surgical instruments or overlapping tissues, remains a critical yet underexplored challenge f

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Seg2Change: Adapting Open-Vocabulary Semantic Segmentation Model for Remote Sensing Change Detection

DGX agent

arXiv:2604.11231v1 Announce Type: new Abstract: Change detection is a fundamental task in remote sensing, aiming to quantify the impacts of human activities and ecological dynamics on land-cover chang

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Select Smarter, Not More: Prompt-Aware Evaluation Scheduling with Submodular Guarantees

DGX agent

arXiv:2604.11328v1 Announce Type: new Abstract: Automatic prompt optimization (APO) hinges on the quality of its evaluation signal, yet scoring every prompt candidate on the full training set is prohi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Self-Evolving LLM Memory Extraction Across Heterogeneous Tasks

DGX agent

arXiv:2604.11610v1 Announce Type: new Abstract: As LLM-based assistants become persistent and personalized, they must extract and retain useful information from past conversations as memory. However,

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Self-supervised Pretraining of Cell Segmentation Models

DGX agent

arXiv:2604.10609v1 Announce Type: new Abstract: Instance segmentation enables the analysis of spatial and temporal properties of cells in microscopy images by identifying the pixels belonging to each

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

DGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Semantic Manipulation Localization

DGX agent

arXiv:2604.10132v1 Announce Type: cross Abstract: Image Manipulation Localization (IML) aims to identify edited regions in an image. However, with the increasing use of modern image editing and genera

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SHANG++: Robust Stochastic Acceleration under Multiplicative Noise

DGX agent

arXiv:2603.09355v1 Announce Type: cross Abstract: Under the multiplicative noise scaling (MNS) condition, original Nesterov acceleration is provably sensitive to noise and may diverge when gradient no

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

SHARE: Social-Humanities AI for Research and Education

DGX agent

arXiv:2604.11152v1 Announce Type: new Abstract: This intermediate technical report introduces the SHARE family of base models and the MIRROR user interface. The SHARE models are the first causal langu

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Shared Emotion Geometry Across Small Language Models: A Cross-Architecture Study of Representation, Behavior, and Methodological Confounds

DGX agent

arXiv:2604.11050v1 Announce Type: cross Abstract: We extract 21-emotion vector sets from twelve small language models (six architectures x base/instruct, 1B-8B parameters) under a unified comprehensio

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Sign Language Recognition in the Age of LLMs

DGX agent

arXiv:2604.11225v1 Announce Type: cross Abstract: Recent Vision Language Models (VLMs) have demonstrated strong performance across a wide range of multimodal reasoning tasks. This raises the question

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units

DGX agent

arXiv:2604.10436v1 Announce Type: new Abstract: Accurate semantic understanding of complex traffic signs-including those with intricate layouts, multi-lingual text, and composite symbols-is critical f

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors

DGX agent

arXiv:2510.17516v4 Announce Type: replace-cross Abstract: Large language model (LLM) simulations of human behavior have the potential to revolutionize the social and behavioral sciences, if and only i

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Simple but Stable, Fast and Safe: Achieve End-to-end Control by High-Fidelity Differentiable Simulation

DGX agent

arXiv:2604.10548v1 Announce Type: new Abstract: Obstacle avoidance is a fundamental vision-based task essential for enabling quadrotors to perform advanced applications. When planning the trajectory,

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis

DGX agent

arXiv:2604.09874v1 Announce Type: new Abstract: Simulating how organized groups (e.g., corporations) make decisions (e.g., responding to a competitor's move) is essential for understanding real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Simulator Adaptation for Sim-to-Real Learning of Legged Locomotion via Proprioceptive Distribution Matching

DGX agent

arXiv:2604.11090v1 Announce Type: new Abstract: Simulation trained legged locomotion policies often exhibit performance loss on hardware due to dynamics discrepancies between the simulator and the rea

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets

DGX agent

arXiv:2604.02460v2 Announce Type: replace Abstract: Recent work reports strong performance from multi-agent LLM systems (MAS), but these gains are often confounded by increased test-time computation.

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

SLM Finetuning for Natural Language to Domain Specific Code Generation in Production

DGX agent

arXiv:2604.09952v1 Announce Type: new Abstract: Many applications today use large language models for code generation; however, production systems have strict latency requirements that can be difficul

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

SMART: When is it Actually Worth Expanding a Speculative Tree?

DGX agent

arXiv:2604.09731v1 Announce Type: cross Abstract: Tree-based speculative decoding accelerates autoregressive generation by verifying a branching tree of draft tokens in a single target-model forward p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation

DGX agent

arXiv:2604.10218v1 Announce Type: new Abstract: Recent self-supervised stereo matching methods have made significant progress. They typically rely on the photometric consistency assumption, which pres

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SmileyLlama: Modifying Large Language Models for Directed Chemical Space Exploration

DGX agent

arXiv:2409.02231v5 Announce Type: replace-cross Abstract: We show that large language model (LLMs) can be transformed via supervised fine-tuning (SFT) of engineered prompts into SmileyLlama for explor

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

SODA: Semi On-Policy Black-Box Distillation for Large Language Models

DGX agent

arXiv:2604.03873v2 Announce Type: replace-cross Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowle

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

DGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Soon, at each gradual improvement level of AI, you will start to see large discrete jumps in ability in economically important areas, becaus…

DGX agent

Soon, at each gradual improvement level of AI, you will start to see large discrete jumps in ability in economically important areas, because the previous AI ability level in some aspect of the job bo

model-releasesethan-mollick--x
14 Apr 2026
Model Releases

Source: Anthropic is preparing to release Claude Opus 4.7, along with a new AI-powered tool for designing websites and presentations, as soon as this week (Stephanie Palazzolo/The Information)

DGX agent

Stephanie Palazzolo / The Information: Source: Anthropic is preparing to release Claude Opus 4.7, along with a new AI-powered tool for designing websites and presentations, as soon as this week — Anth

model-releasestechmeme
14 Apr 2026
Model Releases

Sources: at least two US federal agencies and three congressional committees have reached out to Anthropic to test Claude Mythos, quietly bypassing Trump's ban (Politico)

DGX agent

Politico: Sources: at least two US federal agencies and three congressional committees have reached out to Anthropic to test Claude Mythos, quietly bypassing Trump's ban — The Commerce Department's Ce

model-releasestechmeme
14 Apr 2026
Model Releases

Spatial Competence Benchmark

DGX agent

arXiv:2604.09594v1 Announce Type: new Abstract: Spatial competence is the quality of maintaining a consistent internal representation of an environment and using it to infer discrete structure and pla

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…444445446447448…465
Next →