AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

Beyond Bounded Variance: Variance-Reduced Normalized Methods for Nonconvex Optimization under Blum-Gladyshev Noise

DGX agent

arXiv:2605.15314v1 Announce Type: new Abstract: We study nonconvex stochastic optimization under the Blum-Gladyshev (mathsf{BG}-0) noise model, where the stochastic gradient variance grows quadratical

model-releasesarxiv-cs-lg
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

BiomedAP: A Vision-Informed Dual-Anchor Framework with Gated Cross-Modal Fusion for Robust Medical Vision-Language Adaptation

DGX agent

arXiv:2605.15736v1 Announce Type: cross Abstract: Biomedical Vision--Language Models (VLMs) have shown remarkable promise in few-shot medical diagnosis but face a critical bottleneck: extit{fragility

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Can We Trust AI-Inferred User States. A Psychometric Framework for Validating the Reliability of Users States Classification by LLMs in Operational Environments

DGX agent

arXiv:2605.15734v1 Announce Type: new Abstract: The use of large language models to assess user states in conversational and adaptive systems is based on the assumption that the metrics used for such

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

COPRA: Conditional Parameter Adaptation with Reinforcement Learning for Video Anomaly Detection

DGX agent

arXiv:2605.15325v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown strong performance in video anomaly detection (VAD) while providing interpretable predictions. However, existin

model-releasesarxiv-cs-cv
18 May 2026
Safety

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices

DGX agent

arXiv:2605.15684v1 Announce Type: new Abstract: The Diffusion Transformer (DiT) architecture is the state-of-the-art paradigm for high-fidelity image generation, underpinning models like Stable Diffus

safetyarxiv-cs-cv
18 May 2026
Model Releases

Fully Open Meditron: An Auditable Pipeline for Clinical LLMs

DGX agent

arXiv:2605.16215v1 Announce Type: new Abstract: Clinical decision support systems (CDSS) require scrutable, auditable pipelines that enable rigorous, reproducible validation. Yet current LLM-based CDS

model-releasesarxiv-cs-ai
18 May 2026
Research

GenAI-Driven Approach to RISC-V Supply Chain Exploration

DGX agent

arXiv:2605.15223v1 Announce Type: cross Abstract: This paper presents an LLM-empowered workflow for RISC-V supply chain analysis, integrating Vision-Language Models (VLMs) and Model-Driven Engineering

researcharxiv-cs-ai
18 May 2026
Model Releases

Harnessing Unimodality in Semiparametric Contextual Pricing via Oracle Price Map Learning

DGX agent

arXiv:2605.15411v1 Announce Type: cross Abstract: We study contextual dynamic pricing in a semiparametric scalar-index valuation model where the latent value is v_t=mu_ast(mathsf c_t)+xi_t, with an un

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Hidden in Memory: Sleeper Memory Poisoning in LLM Agents

DGX agent

arXiv:2605.15338v1 Announce Type: cross Abstract: Large language models are increasingly augmented with persistent memory, allowing assistants to store user-specific information across sessions for pe

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

IndicSafe: A Benchmark for Evaluating Multilingual LLM Safety in South Asia

DGX agent

arXiv:2603.17915v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are deployed in multilingual settings, their safety behavior in culturally diverse, low-resource languages rem

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

LoCO: Low-rank Compositional Rotation Fine-tuning

DGX agent

arXiv:2605.15916v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as an critical technique for adapting large-scale foundation models across natural language process

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

MorphoHELM: A Comprehensive Benchmark for Evaluating Representations for Microscopy-Based Morphology Assays

DGX agent

arXiv:2605.15383v1 Announce Type: new Abstract: Microscopy images contain rich information about how cells respond to perturbations, making them essential to applications like drug screening. To quant

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Njord: A Probabilistic Graph Neural Network for Ensemble Ocean Forecasting

DGX agent

arXiv:2605.15470v1 Announce Type: new Abstract: Ocean dynamics are inherently chaotic, yet existing machine learning ocean models produce only deterministic forecasts. We introduce Njord, a probabilis

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control

DGX agent

arXiv:2605.15963v1 Announce Type: new Abstract: Large vision-language models have significantly advanced GUI agents, enabling executable interaction across web, mobile, and desktop interfaces. Yet the

model-releasesarxiv-cs-ai
18 May 2026
Applications

Process-Informed Forecasting of Complex Thermal Dynamics in Pharmaceutical Manufacturing

DGX agent

arXiv:2509.20349v3 Announce Type: replace Abstract: Accurate time-series forecasting for complex physical systems is the backbone of modern industrial monitoring and control, yet deep learning models

applicationsarxiv-cs-lg
18 May 2026
Model Releases

Quantum Feature Pyramid Gating for Seismic Image Segmentation

DGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

DGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

model-releasesarxiv-cs-ai
18 May 2026
Local Ai

Sound Sparks Motion: Audio and Text Tuning for Video Editing

DGX agent

arXiv:2605.15307v1 Announce Type: cross Abstract: Motion-centric video editing remains difficult for large generative video models, which often respond well to appearance changes but struggle to produ

local-aiarxiv-cs-cv
18 May 2026
Model Releases

Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems

DGX agent

arXiv:2601.19923v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into the core of Web-based autonomous agents and complex Web Information Systems, their ability to fait

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SynthRender and IRIS: Open-Source Framework and Dataset for Bidirectional Sim-Real Transfer in Industrial Object Perception

DGX agent

arXiv:2602.21141v2 Announce Type: replace Abstract: Object perception is fundamental for tasks such as robotic material handling and quality inspection. However, modern supervised deep-learning models

model-releasesarxiv-cs-cv
18 May 2026
Agents

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data

DGX agent

arXiv:2509.21465v3 Announce Type: replace Abstract: Tabular foundation models are becoming increasingly popular for low-resource tabular problems. These models make up for small training datasets by p

agentsarxiv-cs-lg
18 May 2026
Model Releases

TokenButler: Token Importance is Predictable

DGX agent

arXiv:2503.07518v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on the Key-Value (KV) Cache to store token history, enabling efficient decoding of tokens. As the KV-Cache g

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

What we announced in streaming AI at Next ‘26

DGX agent

Every device, user, and microservice generates data. Ingesting this data, extracting meaning and insights, and driving business decisions in real time has the potential to deliver transformational bus

model-releasesgoogle-cloud-ai
18 May 2026
Model Releases

G4-MeroMero-31B-uncensored-heretic is Out Now, A finetune of Gemma 4 31B it designed for creative tasks, with KLD of 0.0100 and 15/100 Refusals!

DGX agent

G4-MeroMero-31B-uncensored-heretic is a fine-tuned variant of Gemma 4 31B optimized for creative tasks, featuring low KL divergence (0.0100) and minimal refusals (15/100). The model is designed to be

model-releasesr-ollama
17 May 2026
Model Releases

Introducing Gemini Omni

DGX agent

Gemini Omni is a multimodal AI model that allows users to create content from any input and edit naturally using conversational language . Users can combine images, audio, video, and text as input to

model-releasesgoogle-deepmind
17 May 2026
Model Releases

Recent Developments in LLM Architectures: KV Sharing, mHC, and Compressed Attention

DGX agent

This article discusses recent advancements in large language model architecture design, focusing on key-value (KV) sharing techniques, multi-head cache (mHC) mechanisms, and compressed attention metho

model-releasessebastian-raschka
16 May 2026
Model Releases

ACE-LoRA: Adaptive Orthogonal Decoupling for Continual Image Editing

DGX agent

arXiv:2605.14948v1 Announce Type: new Abstract: State-of-the-art diffusion models often rely on parameter-efficient fine-tuning to perform specialized image editing tasks. However, real-world applicat

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

ArGEnT: Arbitrary Geometry-encoded Transformer for Operator Learning

DGX agent

arXiv:2602.11626v2 Announce Type: replace-cross Abstract: Learning solution operators for systems with complex, varying geometries and parametric physical settings is a central challenge in scientific

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AttnGen: Attention-Guided Saliency Learning for Interpretable Genomic Sequence Classification

DGX agent

arXiv:2605.14073v1 Announce Type: cross Abstract: Deep neural networks have achieved strong performance in genomic sequence classification; however, relating their predictions to biologically meaningf

model-releasesarxiv-cs-ai
15 May 2026
Research

CC-Pan: Channel-wise Compression based Diffusion for Efficient Pan-Sharpening

DGX agent

arXiv:2602.04473v2 Announce Type: replace Abstract: Recently, diffusion models have brought novel insights to pan-sharpening and notably boosted fusion precision. However, most existing models perform

researcharxiv-cs-cv
15 May 2026
Model Releases

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves

DGX agent

arXiv:2605.14068v1 Announce Type: new Abstract: We introduce CurveBench, a benchmark for hierarchical topological reasoning from visual input. CurveBench consists of extbf{756 images} of pairwise non-

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Databricks brings GPT-5.5 to enterprise agent workflows

DGX agent

Databricks has integrated OpenAI's GPT-5.5 model into enterprise agent workflows, enabling organizations to build and deploy AI agents with advanced language capabilities. This partnership leverages D

model-releasesopenai
15 May 2026
Model Releases

Derivation Prompting: A Logic-Based Method for Improving Retrieval-Augmented Generation

DGX agent

arXiv:2605.14053v1 Announce Type: cross Abstract: The application of Large Language Models to Question Answering has shown great promise, but important challenges such as hallucinations and erroneous

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict

DGX agent

arXiv:2605.14473v1 Announce Type: cross Abstract: The Context-Compliance Regime in Retrieval-Augmented Generation (RAG) occurs when retrieved context dominates the final answer even when it conflicts

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Exemplar Partitioning for Mechanistic Interpretability

DGX agent

arXiv:2605.14347v1 Announce Type: new Abstract: We introduce Exemplar Partitioning (EP), an unsupervised method for constructing interpretable feature dictionaries from large language model activation

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

FlowInOne:Unifying Multimodal Generation as Image-in, Image-out Flow Matching

DGX agent

arXiv:2604.06757v2 Announce Type: replace Abstract: Multimodal generation has long been dominated by text-driven pipelines where language dictates vision but cannot reason or create within it. We chal

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Fusion-fission forecasts when AI will shift to undesirable behavior

DGX agent

arXiv:2605.14218v1 Announce Type: new Abstract: The key problem facing ChatGPT-like AI's use across society is that its behavior can shift, unnoticed, from desirable to undesirable -- encouraging self

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GenExam: A Multidisciplinary Text-to-Image Exam

DGX agent

arXiv:2509.14232v5 Announce Type: replace Abstract: Exams are a fundamental test of expert-level intelligence and require integrated understanding, reasoning, and generation. Existing exam-style bench

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

GPart: End-to-End Isometric Fine-Tuning via Global Parameter Partitioning

DGX agent

arXiv:2605.14841v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) has become the dominant paradigm for parameter-efficient fine-tuning (PEFT) of large language models (LLMs). However, its b

model-releasesarxiv-cs-ai
15 May 2026
Safety

GradShield: Alignment Preserving Finetuning

DGX agent

arXiv:2605.14194v1 Announce Type: new Abstract: Large Language Models (LLMs) pose a significant risk of safety misalignment after finetuning, as models can be compromised by both explicitly and implic

safetyarxiv-cs-cl
15 May 2026
Model Releases

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations

DGX agent

arXiv:2605.14498v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly serve as personal assistants and workplace collaborators, where their utility depends on memory systems t

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

How data science teams use Codex

DGX agent

This OpenAI Academy resource explores practical applications of Codex, their AI code generation model, within data science workflows and teams. It likely covers how data scientists leverage Codex to a

model-releasesopenai
15 May 2026
Model Releases

IPR-1: Interactive Physical Reasoner

DGX agent

arXiv:2511.15407v3 Announce Type: replace Abstract: Humans learn by observing, interacting with environments, and internalizing physics and causality. Here, we aim to ask whether an agent can similarl

model-releasesarxiv-cs-ai
15 May 2026
Safety

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis

DGX agent

arXiv:2605.14392v1 Announce Type: new Abstract: We pursue a vision for self-improving language models in which the model does not merely generate problems or traces to imitate, but constructs the envi

safetyarxiv-cs-ai
15 May 2026
Model Releases

LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning

DGX agent

arXiv:2508.06202v2 Announce Type: replace-cross Abstract: Continual Visual Instruction Tuning (CVIT) enables Multimodal Large Language Models (MLLMs) to incrementally learn new tasks over time. Howeve

model-releasesarxiv-cs-ai
15 May 2026
Agents

MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation

DGX agent

arXiv:2602.16898v5 Announce Type: replace-cross Abstract: Task planning for robotic manipulation with large language models (LLMs) is an emerging area. Prior approaches rely on specialized models, fin

agentsarxiv-cs-ai
15 May 2026
Safety

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

DGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

safetyarxiv-cs-ai
15 May 2026
Local Ai

Mistletoe: Stealthy Acceleration-Collapse Attacks on Speculative Decoding

DGX agent

arXiv:2605.14005v1 Announce Type: new Abstract: Speculative decoding has become a widely adopted technique for accelerating large language model (LLM) inference by drafting multiple candidate tokens a

local-aiarxiv-cs-cl
15 May 2026
← Previous
1…466467468469470…1371
Next →