AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,338
  • Agents7,718
  • Applications5,508
  • Concepts5
  • Hardware1,906
  • Industry6,195
  • Local Ai5,052
  • Model Releases24,549
  • Research20,616
  • Safety13,637
  • Syntheses17
  • Tools1,678
  • Tutorials3,457

Source
HumanDGX agent

Content type
AllBlog
90,338Total entries
1Added by human
90,337Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,187 results
Local Ai

Correct me if I’m wrong: Ollama can’t fine tune like Unsloth Studio

DGX agent

Ollama is a local inference engine designed for running pre-built LLMs on your own machine, and it does not include fine-tuning capabilities — this distinction is correct. Unsloth (and its Unsloth Stu

local-air-ollama
14 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing

DGX agent

arXiv:2506.18438v2 Announce Type: replace Abstract: Editing natural images using textual descriptions in text-to-image diffusion models remains a significant challenge, particularly in achieving consi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

CricBench: A Multilingual Benchmark for Evaluating LLMs in Cricket Analytics

DGX agent

arXiv:2512.21877v3 Announce Type: replace-cross Abstract: Cricket is the second most popular sport worldwide, with billions of fans seeking advanced statistical insights unavailable through standard w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

DecepGPT: Schema-Driven Deception Detection with Multicultural Datasets and Robust Multimodal Learning

DGX agent

arXiv:2603.23916v2 Announce Type: replace-cross Abstract: Multimodal deception detection aims to identify deceptive behavior by analyzing audiovisual cues for forensics and security. In these high-sta

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Defending against Backdoor Attacks via Module Switching

DGX agent

arXiv:2504.05902v2 Announce Type: replace-cross Abstract: Backdoor attacks pose a serious threat to deep neural networks (DNNs), allowing adversaries to implant triggers for hidden behaviors in infere

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Delta Rectified Flow Sampling for Text-to-Image Editing

DGX agent

arXiv:2509.05342v3 Announce Type: replace Abstract: We propose Delta Rectified Flow Sampling (DRFS), a novel inversion-free, path-aware editing framework within rectified flow models for text-to-image

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

DERM-3R: A Resource-Efficient Multimodal Agents Framework for Dermatologic Diagnosis and Treatment in Real-World Clinical Settings

DGX agent

arXiv:2604.09596v1 Announce Type: new Abstract: Dermatologic diseases impose a large and growing global burden, affecting billions and substantially reducing quality of life. While modern therapies ca

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

DiningBench: A Hierarchical Multi-view Benchmark for Perception and Reasoning in the Dietary Domain

DGX agent

arXiv:2604.10425v1 Announce Type: new Abstract: Recent advancements in Vision-Language Models (VLMs) have revolutionized general visual understanding. However, their application in the food domain rem

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation

DGX agent

arXiv:2503.16771v3 Announce Type: replace-cross Abstract: As Large Language Models for Code (LM4Code) become integral to software engineering, establishing trust in their output becomes critical. Howe

local-aiarxiv-cs-lg
14 Apr 2026
Safety

Enhancing Fine-Grained Spatial Grounding in 3D CT Report Generation via Discriminative Guidance

DGX agent

arXiv:2604.10437v1 Announce Type: new Abstract: Vision--language models (VLMs) for radiology report generation (RRG) can produce long-form chest CT reports from volumetric scans and show strong potent

safetyarxiv-cs-cv
14 Apr 2026
Local Ai

Ernie Image Turbo is Capable of ...

DGX agent

ERNIE Image Turbo is an open text-to-image generation model developed by Baidu's ERNIE-Image team, serving as the distilled release of the full ERNIE-Image model and built on a single-stream Diffusion

local-air-stablediffusion
14 Apr 2026
Research

Evaluating Visual Prompts with Eye-Tracking Data for MLLM-Based Human Activity Recognition

DGX agent

arXiv:2604.09585v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as foundation models for IoT applications such as human activity recognition (HAR). However, directly applyi

researcharxiv-cs-ai
14 Apr 2026
Model Releases

GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs

DGX agent

arXiv:2511.11653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for passage reranking in information retrieval, leveraging their superior reasonin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

DGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

model-releasesharrison-chase--x
14 Apr 2026
Tutorials

HiddenObjects: Scalable Diffusion-Distilled Spatial Priors for Object Placement

DGX agent

arXiv:2604.10675v1 Announce Type: new Abstract: We propose a method to learn explicit, class-conditioned spatial priors for object placement in natural scenes by distilling the implicit placement know

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

DGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Human vs. Machine Deception: Distinguishing AI-Generated and Human-Written Fake News Using Ensemble Learning

DGX agent

arXiv:2604.09960v1 Announce Type: new Abstract: The rapid adoption of large language models has introduced a new class of AI-generated fake news that coexists with traditional human-written misinforma

researcharxiv-cs-cl
14 Apr 2026
Model Releases

Intent-aligned Formal Specification Synthesis via Traceable Refinement

DGX agent

arXiv:2604.10392v1 Announce Type: cross Abstract: Large language models are increasingly used to generate code from natural language, but ensuring correctness remains challenging. Formal verification

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs

DGX agent

arXiv:2604.10403v1 Announce Type: new Abstract: We address jailbreaks, backdoors, and unlearning for large language models (LLMs). Unlike prior work, which trains LLMs based on their actions when give

safetyarxiv-cs-lg
14 Apr 2026
Tutorials

Learning and Enforcing Context-Sensitive Control for LLMs

DGX agent

arXiv:2604.10667v1 Announce Type: cross Abstract: Controlling the output of Large Language Models (LLMs) through context-sensitive constraints has emerged as a promising approach to overcome the limit

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

Learning to Adapt: In-Context Learning Beyond Stationarity

DGX agent

arXiv:2604.10946v1 Announce Type: new Abstract: Transformer models have become foundational across a wide range of scientific and engineering domains due to their strong empirical performance. A key c

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

LLMs Should Incorporate Explicit Mechanisms for Human Empathy

DGX agent

arXiv:2604.10557v1 Announce Type: cross Abstract: This paper argues that Large Language Models (LLMs) should incorporate explicit mechanisms for human empathy. As LLMs become increasingly deployed in

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

DGX agent

arXiv:2601.14706v3 Announce Type: replace Abstract: In this paper, we present LookBench (We use the term 'look' to reflect retrieval that mirrors how people shop -- finding the exact item, a close sub

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LRD-Net: A Lightweight Real-Centered Detection Network for Cross-Domain Face Forgery Detection

DGX agent

arXiv:2604.10862v1 Announce Type: new Abstract: The rapid advancement of diffusion-based generative models has made face forgery detection a critical challenge in digital forensics. Current detection

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

MARS: Unleashing the Power of Speculative Decoding via Margin-Aware Verification

DGX agent

arXiv:2601.15498v2 Announce Type: replace Abstract: Speculative Decoding (SD) accelerates autoregressive large language model (LLM) inference by decoupling generation and verification. While recent me

local-aiarxiv-cs-lg
14 Apr 2026
Model Releases

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

DGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Min-k Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics

DGX agent

arXiv:2604.11012v1 Announce Type: new Abstract: The quality of text generated by large language models depends critically on the decoding sampling strategy. While mainstream methods such as Top-k, Top

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark

DGX agent

arXiv:2604.10755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have advanced clinical tasks for common conditions, but their performance on rare diseases remains largely unte

model-releasesarxiv-cs-cv
14 Apr 2026
Hardware

MoBo, CPU, RAM suggestion | I'm terrible at guessing hardware | no gpu

DGX agent

This Reddit thread from r/ollama features a user seeking community recommendations for motherboard, CPU, and RAM components to build a system optimized for running Ollama locally without a dedicated G

hardwarer-ollama
14 Apr 2026
Model Releases

Multi-modal, multi-scale representation learning for satellite imagery analysis just needs a good ALiBi

DGX agent

arXiv:2604.10347v1 Announce Type: new Abstract: Vision foundation models have been shown to be effective at processing satellite imagery into representations fit for downstream tasks, however, creatin

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

Multimodal Diffusion Forcing for Forceful Manipulation

DGX agent

arXiv:2511.04812v2 Announce Type: replace-cross Abstract: Given a dataset of expert trajectories, standard imitation learning approaches typically learn a direct mapping from observations (e.g., RGB i

tutorialsarxiv-cs-ai
14 Apr 2026
Research

Near OOD Detection for Vision-Language Prompt Learning with Contrastive Logit Score

DGX agent

arXiv:2405.16091v2 Announce Type: replace Abstract: Prompt learning has emerged as an efficient and effective method for fine-tuning vision-language models such as CLIP. While many studies have explor

researcharxiv-cs-cv
14 Apr 2026
Applications

Non-stationary Diffusion For Probabilistic Time Series Forecasting

DGX agent

arXiv:2505.04278v3 Announce Type: replace-cross Abstract: Due to the dynamics of underlying physics and external influences, the uncertainty of time series often varies over time. However, existing De

applicationsarxiv-cs-ai
14 Apr 2026
Research

On The Application of Linear Attention in Multimodal Transformers

DGX agent

arXiv:2604.10064v1 Announce Type: new Abstract: Multimodal Transformers serve as the backbone for state-of-the-art vision-language models, yet their quadratic attention complexity remains a critical b

researcharxiv-cs-cv
14 Apr 2026
Local Ai

Pair2Scene: Learning Local Object Relations for Procedural Scene Generation

DGX agent

arXiv:2604.11808v1 Announce Type: new Abstract: Generating high-fidelity 3D indoor scenes remains a significant challenge due to data scarcity and the complexity of modeling intricate spatial relation

local-aiarxiv-cs-cv
14 Apr 2026
Safety

PEMANT: Persona-Enriched Multi-Agent Negotiation for Travel

DGX agent

arXiv:2604.10475v1 Announce Type: new Abstract: Modeling household-level trip generation is fundamental to accurate demand forecasting, traffic flow estimation, and urban system planning. Existing stu

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Persona Non Grata: Single-Method Safety Evaluation Is Incomplete for Persona-Imbued LLMs

DGX agent

arXiv:2604.11120v1 Announce Type: new Abstract: Personality imbuing customizes LLM behavior, but safety evaluations almost always study prompt-based personas alone. We show this is incomplete: prompti

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Query Lower Bounds for Diffusion Sampling

DGX agent

arXiv:2604.10857v1 Announce Type: cross Abstract: Diffusion models generate samples by iteratively querying learned score estimates. A rapidly growing literature focuses on accelerating sampling by mi

researcharxiv-cs-ai
14 Apr 2026
Applications

Resisting Humanization: Ethical Front-End Design Choices in AI for Sensitive Contexts

DGX agent

arXiv:2603.24853v2 Announce Type: replace Abstract: Ethical debates in AI have primarily focused on back-end issues such as data governance, model training, and algorithmic decision-making. Less atten

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies

DGX agent

arXiv:2604.09860v1 Announce Type: cross Abstract: The pursuit of general-purpose robotics has yielded impressive foundation models, yet simulation-based benchmarking remains a bottleneck due to rapid

model-releasesarxiv-cs-ai
14 Apr 2026
Research

S4M: 4-points to Segment Anything

DGX agent

arXiv:2503.05534v3 Announce Type: replace Abstract: Purpose: The Segment Anything Model (SAM) promises to ease the annotation bottleneck in medical segmentation, but overlapping anatomy and blurred bo

researcharxiv-cs-cv
14 Apr 2026
Safety

SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering

DGX agent

arXiv:2508.11290v3 Announce Type: replace Abstract: LLMs increasingly exhibit over-refusal behavior, where safety mechanisms cause models to reject benign instructions that seemingly resemble harmful

safetyarxiv-cs-cl
14 Apr 2026
Hardware

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

DGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

hardwarearxiv-cs-cv
14 Apr 2026
Model Releases

SciPredict: Can LLMs Predict the Outcomes of Scientific Experiments in Natural Sciences?

DGX agent

arXiv:2604.10718v1 Announce Type: new Abstract: Accelerating scientific discovery requires the identification of which experiments would yield the best outcomes before committing resources to costly p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding

DGX agent

arXiv:2604.11244v1 Announce Type: new Abstract: Advances in Multimodal Large Language Models (MLLMs) are transforming video captioning from a descriptive endpoint into a semantic interface for both vi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

SpecMoE: A Fast and Efficient Mixture-of-Experts Inference via Self-Assisted Speculative Decoding

DGX agent

arXiv:2604.10152v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs)

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SPEED-Bench: A Unified and Diverse Benchmark for Speculative Decoding

DGX agent

arXiv:2604.09557v1 Announce Type: cross Abstract: Speculative Decoding (SD) has emerged as a critical technique for accelerating Large Language Model (LLM) inference. Unlike deterministic system optim

model-releasesarxiv-cs-ai
14 Apr 2026
Research

StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs

DGX agent

arXiv:2509.22220v2 Announce Type: replace Abstract: Prevalent semantic speech tokenizers, designed to capture linguistic content, are surprisingly fragile. We find they are not robust to meaning-irrel

researcharxiv-cs-cl
14 Apr 2026
← Previous
1…479480481482483…1359
Next →