AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlog
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,128 results
Model Releases

Hidden in Memory: Sleeper Memory Poisoning in LLM Agents

DGX agent

arXiv:2605.15338v1 Announce Type: cross Abstract: Large language models are increasingly augmented with persistent memory, allowing assistants to store user-specific information across sessions for pe

model-releasesarxiv-cs-ai
18 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia

DGX agent

arXiv:2603.17915v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are deployed in multilingual settings, their safety behavior in culturally diverse, low-resource languages rem

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

LoCO: Low-rank Compositional Rotation Fine-tuning

DGX agent

arXiv:2605.15916v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as an critical technique for adapting large-scale foundation models across natural language process

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays

DGX agent

arXiv:2605.15383v1 Announce Type: new Abstract: Microscopy images contain rich information about how cells respond to perturbations, making them essential to applications like drug screening. To quant

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Njord: A Probabilistic Graph Neural Network for Ensemble Ocean Forecasting

DGX agent

arXiv:2605.15470v1 Announce Type: new Abstract: Ocean dynamics are inherently chaotic, yet existing machine learning ocean models produce only deterministic forecasts. We introduce Njord, a probabilis

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control

DGX agent

arXiv:2605.15963v1 Announce Type: new Abstract: Large vision-language models have significantly advanced GUI agents, enabling executable interaction across web, mobile, and desktop interfaces. Yet the

model-releasesarxiv-cs-ai
18 May 2026
Applications

Process-Informed Forecasting of Complex Thermal Dynamics in Pharmaceutical Manufacturing

DGX agent

arXiv:2509.20349v3 Announce Type: replace Abstract: Accurate time-series forecasting for complex physical systems is the backbone of modern industrial monitoring and control, yet deep learning models

applicationsarxiv-cs-lg
18 May 2026
Model Releases

Quantum Feature Pyramid Gating for Seismic Image Segmentation

DGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

DGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

model-releasesarxiv-cs-ai
18 May 2026
Local Ai

Sound Sparks Motion: Audio and Text Tuning for Video Editing

DGX agent

arXiv:2605.15307v1 Announce Type: cross Abstract: Motion-centric video editing remains difficult for large generative video models, which often respond well to appearance changes but struggle to produ

local-aiarxiv-cs-cv
18 May 2026
Model Releases

Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems

DGX agent

arXiv:2601.19923v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into the core of Web-based autonomous agents and complex Web Information Systems, their ability to fait

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception

DGX agent

arXiv:2602.21141v2 Announce Type: replace Abstract: Object perception is fundamental for tasks such as robotic material handling and quality inspection. However, modern supervised deep-learning models

model-releasesarxiv-cs-cv
18 May 2026
Agents

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data

DGX agent

arXiv:2509.21465v3 Announce Type: replace Abstract: Tabular foundation models are becoming increasingly popular for low-resource tabular problems. These models make up for small training datasets by p

agentsarxiv-cs-lg
18 May 2026
Model Releases

TokenButler: Token Importance is Predictable

DGX agent

arXiv:2503.07518v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache g

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

What we announced in streaming AI at Next ‘26

DGX agent

Every device, user, and microservice generates data. Ingesting this data, extracting meaning and insights, and driving business decisions in real time has the potential to deliver transformational bus

model-releasesgoogle-cloud-ai
18 May 2026
Model Releases

G4-MeroMero-31B-uncensored-heretic is Out Now, A finetune of Gemma 4 31B it designed for creative tasks, with KLD of 0.0100 and 15/100 Refusals!

DGX agent

G4-MeroMero-31B-uncensored-heretic is a fine-tuned variant of Gemma 4 31B optimized for creative tasks, featuring low KL divergence (0.0100) and minimal refusals (15/100). The model is designed to be

model-releasesr-ollama
17 May 2026
Model Releases

Introducing Gemini Omni

DGX agent

Gemini Omni is a multimodal AI model that allows users to create content from any input and edit naturally using conversational language . Users can combine images, audio, video, and text as input to

model-releasesgoogle-deepmind
17 May 2026
Model Releases

Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention

DGX agent

This article discusses recent advancements in large language model architecture design, focusing on key-value (KV) sharing techniques, multi-head cache (mHC) mechanisms, and compressed attention metho

model-releasessebastian-raschka
16 May 2026
Model Releases

ACE-LoRA: Adaptive Orthogonal Decoupling for Continual Image Editing

DGX agent

arXiv:2605.14948v1 Announce Type: new Abstract: State-of-the-art diffusion models often rely on parameter-efficient fine-tuning to perform specialized image editing tasks. However, real-world applicat

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

DGX agent

arXiv:2602.11626v2 Announce Type: replace-cross Abstract: Learning solution operators for systems with complex, varying geometries and parametric physical settings is a central challenge in scientific

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AttnGen: Attention-Guided Saliency Learning for Interpretable Genomic Sequence Classification

DGX agent

arXiv:2605.14073v1 Announce Type: cross Abstract: Deep neural networks have achieved strong performance in genomic sequence classification; however, relating their predictions to biologically meaningf

model-releasesarxiv-cs-ai
15 May 2026
Research

CC-Pan: Channel-wise Compression based Diffusion for Efficient Pan-Sharpening

DGX agent

arXiv:2602.04473v2 Announce Type: replace Abstract: Recently, diffusion models have brought novel insights to pan-sharpening and notably boosted fusion precision. However, most existing models perform

researcharxiv-cs-cv
15 May 2026
Model Releases

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

DGX agent

arXiv:2605.14068v1 Announce Type: new Abstract: We introduce CurveBench, a benchmark for hierarchical topological reasoning from visual input. CurveBench consists of extbf{756 images} of pairwise non-

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Databricks brings GPT-5.5 to enterprise agent workflows

DGX agent

Databricks has integrated OpenAI's GPT-5.5 model into enterprise agent workflows, enabling organizations to build and deploy AI agents with advanced language capabilities. This partnership leverages D

model-releasesopenai
15 May 2026
Model Releases

Derivation Prompting: A Logic-Based Method for Improving Retrieval-Augmented Generation

DGX agent

arXiv:2605.14053v1 Announce Type: cross Abstract: The application of Large Language Models to Question Answering has shown great promise, but important challenges such as hallucinations and erroneous

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict

DGX agent

arXiv:2605.14473v1 Announce Type: cross Abstract: The Context-Compliance Regime in Retrieval-Augmented Generation (RAG) occurs when retrieved context dominates the final answer even when it conflicts

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Exemplar Partitioning for Mechanistic Interpretability

DGX agent

arXiv:2605.14347v1 Announce Type: new Abstract: We introduce Exemplar Partitioning (EP), an unsupervised method for constructing interpretable feature dictionaries from large language model activation

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

FlowInOne:Unifying Multimodal Generation as Image-in, Image-out Flow Matching

DGX agent

arXiv:2604.06757v2 Announce Type: replace Abstract: Multimodal generation has long been dominated by text-driven pipelines where language dictates vision but cannot reason or create within it. We chal

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Fusion-fission forecasts when AI will shift to undesirable behavior

DGX agent

arXiv:2605.14218v1 Announce Type: new Abstract: The key problem facing ChatGPT-like AI's use across society is that its behavior can shift, unnoticed, from desirable to undesirable -- encouraging self

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GenExam: A Multidisciplinary Text-to-Image Exam

DGX agent

arXiv:2509.14232v5 Announce Type: replace Abstract: Exams are a fundamental test of expert-level intelligence and require integrated understanding, reasoning, and generation. Existing exam-style bench

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

GPart: End-to-End Isometric Fine-Tuning via Global Parameter Partitioning

DGX agent

arXiv:2605.14841v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become the dominant paradigm for parameter-efficient fine-tuning (PEFT) of large language models (LLMs). However, its b

model-releasesarxiv-cs-ai
15 May 2026
Safety

GradShield: Alignment Preserving Finetuning

DGX agent

arXiv:2605.14194v1 Announce Type: new Abstract: Large Language Models (LLMs) pose a significant risk of safety misalignment after finetuning, as models can be compromised by both explicitly and implic

safetyarxiv-cs-cl
15 May 2026
Model Releases

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations

DGX agent

arXiv:2605.14498v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly serve as personal assistants and workplace collaborators, where their utility depends on memory systems t

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

How data science teams use Codex

DGX agent

This OpenAI Academy resource explores practical applications of Codex, their AI code generation model, within data science workflows and teams. It likely covers how data scientists leverage Codex to a

model-releasesopenai
15 May 2026
Model Releases

IPR-1: Interactive Physical Reasoner

DGX agent

arXiv:2511.15407v3 Announce Type: replace Abstract: Humans learn by observing, interacting with environments, and internalizing physics and causality. Here, we aim to ask whether an agent can similarl

model-releasesarxiv-cs-ai
15 May 2026
Safety

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis

DGX agent

arXiv:2605.14392v1 Announce Type: new Abstract: We pursue a vision for self-improving language models in which the model does not merely generate problems or traces to imitate, but constructs the envi

safetyarxiv-cs-ai
15 May 2026
Model Releases

LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning

DGX agent

arXiv:2508.06202v2 Announce Type: replace-cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models (MLLMs) to incrementally learn new tasks over time. Howeve

model-releasesarxiv-cs-ai
15 May 2026
Agents

MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation

DGX agent

arXiv:2602.16898v5 Announce Type: replace-cross Abstract: Task planning for robotic manipulation with large language models (LLMs) is an emerging area. Prior approaches rely on specialized models, fin

agentsarxiv-cs-ai
15 May 2026
Safety

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

DGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

safetyarxiv-cs-ai
15 May 2026
Local Ai

Mistletoe: Stealthy Acceleration-Collapse Attacks on Speculative Decoding

DGX agent

arXiv:2605.14005v1 Announce Type: new Abstract: Speculative decoding has become a widely adopted technique for accelerating large language model (LLM) inference by drafting multiple candidate tokens a

local-aiarxiv-cs-cl
15 May 2026
Model Releases

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation

DGX agent

arXiv:2605.13857v1 Announce Type: cross Abstract: The creation of cinematic-quality animal effects necessitates the precise modeling of muscle and fur dynamics, a process that remains both labor-inten

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

MUON+: Towards More Effective Muon via One Additional Normalization Step for LLM Pre-training

DGX agent

arXiv:2602.21545v3 Announce Type: replace Abstract: Muon has recently emerged as a strong optimizer for large language model pre-training, orthogonalizing the momentum matrix via Newton--Schulz polar

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Neural Signals Generate Clinical Notes in the Wild

DGX agent

arXiv:2601.22197v3 Announce Type: replace-cross Abstract: Generating clinical reports that summarize abnormal patterns, diagnostic findings, and clinical interpretations from long-term EEG recordings

model-releasesarxiv-cs-ai
15 May 2026
Research

RoSHAP: A Distributional Framework and Robust Metric for Stable Feature Attribution

DGX agent

arXiv:2605.15154v1 Announce Type: cross Abstract: Feature attribution analysis is critical for interpreting machine learning models and supporting reliable data-driven decisions. However, feature attr

researcharxiv-cs-lg
15 May 2026
Model Releases

SceneFunRI: Reasoning the Invisible for Task-Driven Functional Object Localization

DGX agent

arXiv:2605.14704v1 Announce Type: cross Abstract: In real-world scenes, target objects may reside in regions that are not visible. While humans can often infer the locations of occluded objects from c

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

SCOOTER: A Human Evaluation Framework for Unrestricted Adversarial Examples

DGX agent

arXiv:2507.07776v3 Announce Type: replace Abstract: Unrestricted adversarial attacks aim to fool computer vision models without being constrained by ell_p-norm bounds to remain imperceptible to humans

model-releasesarxiv-cs-cv
15 May 2026
Tutorials

Support Before Frequency in Discrete Diffusion

DGX agent

arXiv:2605.13999v1 Announce Type: new Abstract: Discrete diffusion models are increasingly competitive for language modeling, yet it remains unclear how their denoising objectives organize learning. A

tutorialsarxiv-cs-lg
15 May 2026
Model Releases

SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades

DGX agent

arXiv:2605.14415v1 Announce Type: cross Abstract: Coding agents powered by large language models are increasingly expected to perform realistic software maintenance tasks beyond isolated issue resolut

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…461462463464465…1357
Next →