AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Research

Refined Differentially Private Linear Regression via Extension of a Free Lunch Result

DGX agent

arXiv:2604.11820v1 Announce Type: cross Abstract: As data-privacy regulations tighten and statistical models are increasingly deployed on sensitive human-sourced data, privacy-preserving linear regres

researcharxiv-cs-lg
15 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

RPG-SAM: Reliability-Weighted Prototypes and Geometric Adaptive Threshold Selection for Training-Free One-Shot Polyp Segmentation

DGX agent

arXiv:2603.07436v2 Announce Type: replace Abstract: Training-free one-shot segmentation offers a scalable alternative to expert annotations where knowledge is often transferred from support images and

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Socrates Loss: Unifying Confidence Calibration and Classification by Leveraging the Unknown

DGX agent

arXiv:2604.12245v1 Announce Type: cross Abstract: Deep neural networks, despite their high accuracy, often exhibit poor confidence calibration, limiting their reliability in high-stakes applications.

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

StoryScope: Investigating idiosyncrasies in AI fiction

DGX agent

arXiv:2604.03136v4 Announce Type: replace Abstract: As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluate

model-releasesarxiv-cs-cl
15 Apr 2026
Model Releases

Subspace-Guided Feature Reconstruction for Unsupervised Anomaly Localization

DGX agent

arXiv:2309.13904v3 Announce Type: replace Abstract: Unsupervised anomaly localization aims to identify anomalous regions that deviate from normal sample patterns. Most recent methods perform feature m

model-releasesarxiv-cs-cv
15 Apr 2026
Safety

TEMPLATEFUZZ: Fine-Grained Chat Template Fuzzing for Jailbreaking and Red Teaming LLMs

DGX agent

arXiv:2604.12232v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed across diverse domains, yet their vulnerability to jailbreak attacks, where adversarial inputs

safetyarxiv-cs-ai
15 Apr 2026
Industry

The ‘Goldilocks zone’: How the AI factory ends the cycle of rebuilding pipelines from scratch

DGX agent

The AI bottleneck isn’t the model — it’s everything that has to happen to the data before the model can even touch it. Conversational analytics is now emerging as the bridge to turn already-curated da

industrysiliconangle
15 Apr 2026
Agents

Thermodynamic Liquid Manifold Networks: Physics-Bounded Deep Learning for Solar Forecasting in Autonomous Off-Grid Microgrids

DGX agent

arXiv:2604.11909v1 Announce Type: cross Abstract: The stable operation of autonomous off-grid photovoltaic systems requires solar forecasting algorithms that respect atmospheric thermodynamics. Contem

agentsarxiv-cs-ai
15 Apr 2026
Safety

Token-Level Policy Optimization: Linking Group-Level Rewards to Token-Level Aggregation via Sequence-Level Likelihood

DGX agent

arXiv:2604.12736v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has significantly advanced the reasoning ability of large language models (LLMs), particularly in their mathem

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

Towards Long-horizon Agentic Multimodal Search

DGX agent

arXiv:2604.12890v1 Announce Type: cross Abstract: Multimodal deep search agents have shown great potential in solving complex tasks by iteratively collecting textual and visual evidence. However, mana

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Towards Realistic and Consistent Orbital Video Generation via 3D Foundation Priors

DGX agent

arXiv:2604.12309v1 Announce Type: new Abstract: We present a novel method for generating geometrically realistic and consistent orbital videos from a single image of an object. Existing video generati

researcharxiv-cs-cv
15 Apr 2026
Model Releases

Tree Learning: A Multi-Skill Continual Learning Framework for Humanoid Robots

DGX agent

arXiv:2604.12909v1 Announce Type: new Abstract: As reinforcement learning for humanoid robots evolves from single-task to multi-skill paradigms, efficiently expanding new skills while avoiding catastr

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Visual Preference Optimization with Rubric Rewards

DGX agent

arXiv:2604.13029v1 Announce Type: cross Abstract: The effectiveness of Direct Preference Optimization (DPO) depends on preference data that reflect the quality differences that matter in multimodal ta

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

VPTracker: Global Vision-Language Tracking via Visual Prompt

DGX agent

arXiv:2512.22799v2 Announce Type: replace Abstract: Vision-Language Tracking aims to continuously localize objects described by a visual template and a language description. Existing methods, however,

local-aiarxiv-cs-cv
15 Apr 2026
Model Releases

X-VC: Zero-shot Streaming Voice Conversion in Codec Space

DGX agent

arXiv:2604.12456v1 Announce Type: cross Abstract: Zero-shot voice conversion (VC) aims to convert a source utterance into the voice of an unseen target speaker while preserving its linguistic content.

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ACE-Bench: A Lightweight Benchmark for Evaluating Azure SDK Usage Correctness

DGX agent

arXiv:2604.09564v1 Announce Type: cross Abstract: We present ACE-Bench (Azure SDK Coding Evaluation Benchmark), an execution-free benchmark that provides fast, reproducible pass or fail signals for wh

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Adversarial Video Promotion Against Text-to-Video Retrieval

DGX agent

arXiv:2508.06964v3 Announce Type: replace Abstract: Thanks to the development of cross-modal models, text-to-video retrieval (T2VR) is advancing rapidly, but its robustness remains largely unexamined.

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Agent^2 RL-Bench: Can LLM Agents Engineer Agentic RL Post-Training?

DGX agent

arXiv:2604.10547v1 Announce Type: new Abstract: We introduce Agent^2 RL-Bench, a benchmark for evaluating agentic RL post-training -- whether LLM agents can autonomously design, implement, and run com

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

Agents in Ollama and Langflow

DGX agent

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

local-air-ollama
14 Apr 2026
Research

Are Pretrained Image Matchers Good Enough for SAR-Optical Satellite Registration?

DGX agent

arXiv:2604.10217v1 Announce Type: new Abstract: Cross-modal optical-SAR (Synthetic Aperture Radar) registration is a bottleneck for disaster-response via remote sensing, yet modern image matchers are

researcharxiv-cs-cv
14 Apr 2026
Local Ai

At FullTilt: Real-Time Open-Set 3D Macromolecule Detection Directly from Tilted 2D Projections

DGX agent

arXiv:2604.10766v1 Announce Type: new Abstract: Open-set 3D macromolecule detection in cryogenic electron tomography eliminates the need for target-specific model retraining. However, strict VRAM cons

local-aiarxiv-cs-cv
14 Apr 2026
Safety

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels

DGX agent

arXiv:2604.10367v1 Announce Type: new Abstract: Audio-driven human video generation has achieved remarkable success in monologue scenarios, largely driven by advancements in powerful video generation

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Beyond the Beep: Scalable Collision Anticipation and Real-Time Explainability with BADAS-2.0

DGX agent

arXiv:2604.05767v2 Announce Type: replace-cross Abstract: We present BADAS-2.0, the second generation of our collision anticipation system, building on BADAS-1.0, which showed that fine-tuning V-JEPA2

model-releasesarxiv-cs-cl
14 Apr 2026
Research

BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation

DGX agent

arXiv:2603.14410v3 Announce Type: replace Abstract: Generating long-form linear fiction from open-ended themes remains a major challenge for large language models, which frequently fail to guarantee g

researcharxiv-cs-cl
14 Apr 2026
Model Releases

BITS Pilani at SemEval-2026 Task 9: Structured Supervised Fine-Tuning with DPO Refinement for Polarization Detection

DGX agent

arXiv:2604.11121v1 Announce Type: new Abstract: The POLAR SemEval-2026 Shared Task aims to detect online polarization and focuses on the classification and identification of multilingual, multicultura

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

BlasBench: An Open Benchmark for Irish Speech Recognition

DGX agent

arXiv:2604.10736v1 Announce Type: new Abstract: No open Irish-specific benchmark compares end-user ASR systems under a shared Irish-aware evaluation protocol. To solve this, we release BlasBench, an o

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

DGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Bottleneck Tokens for Unified Multimodal Retrieval

DGX agent

arXiv:2604.11095v1 Announce Type: cross Abstract: Adapting decoder-only multimodal large language models (MLLMs) for unified multimodal retrieval faces two structural gaps. First, existing methods rel

researcharxiv-cs-ai
14 Apr 2026
Research

Catalog-Native LLM: Speaking Item-ID Dialect with Less Entanglement for Recommendation

DGX agent

arXiv:2510.05125v2 Announce Type: replace Abstract: While collaborative filtering delivers predictive accuracy and efficiency, and Large Language Models (LLMs) enable expressive and generalizable reas

researcharxiv-cs-cl
14 Apr 2026
Model Releases

CocoaBench: Evaluating Unified Digital Agents in the Wild

DGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CoFusion: Multispectral and Hyperspectral Image Fusion via Spectral Coordinate Attention

DGX agent

arXiv:2604.10584v1 Announce Type: new Abstract: Multispectral and Hyperspectral Image Fusion (MHIF) aims to reconstruct high-resolution images by integrating low-resolution hyperspectral images (LRHSI

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

COMPOSITE-Stem

DGX agent

arXiv:2604.09836v1 Announce Type: new Abstract: AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Context-Aware Semantic Segmentation via Stage-Wise Attention

DGX agent

arXiv:2601.11310v2 Announce Type: replace Abstract: Semantic ultra-high-resolution (UHR) image segmentation is essential in remote sensing applications such as aerial mapping and environmental monitor

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training

DGX agent

arXiv:2507.15640v2 Announce Type: replace-cross Abstract: Continual pre-training on small-scale task-specific data is an effective method for improving large language models in new target fields, yet

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

DDO-RM for LLM Preference Optimization: A Minimal Held-Out Benchmark against DPO

DGX agent

arXiv:2604.11119v1 Announce Type: cross Abstract: This paper reorganizes the current manuscript around the DPO versus DDO-RM preference-optimization project and focuses on two parts: the algorithmic v

model-releasesarxiv-cs-lg
14 Apr 2026
Research

DeepSketcher: Internalizing Visual Manipulation for Multimodal Reasoning

DGX agent

arXiv:2509.25866v2 Announce Type: replace Abstract: The 'thinking with images' paradigm represents a pivotal shift in the reasoning of Vision Language Models (VLMs), moving from text-dominant chain-of

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Differentially Private Verification of Distribution Properties

DGX agent

arXiv:2604.10819v1 Announce Type: cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of

model-releasesarxiv-cs-lg
14 Apr 2026
Safety

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

DGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

safetyarxiv-cs-ai
14 Apr 2026
Research

Diffusion-CAM: Faithful Visual Explanations for dMLLMs

DGX agent

arXiv:2604.11005v1 Announce Type: new Abstract: While diffusion Multimodal Large Language Models (dMLLMs) have recently achieved remarkable strides in multimodal generation, the development of interpr

researcharxiv-cs-ai
14 Apr 2026
Research

Discourse Diversity in Multi-Turn Empathic Dialogue

DGX agent

arXiv:2604.11742v1 Announce Type: cross Abstract: Large language models (LLMs) produce responses rated as highly empathic in single-turn settings (Ayers et al., 2023; Lee et al., 2024), yet they are a

researcharxiv-cs-ai
14 Apr 2026
Safety

Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations

DGX agent

arXiv:2604.11322v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive capabilities in utilizing external tools. In practice, however, LLMs are often exposed to to

safetyarxiv-cs-ai
14 Apr 2026
Applications

EDUMATH: Generating Standards-aligned Educational Math Word Problems

DGX agent

arXiv:2510.06965v2 Announce Type: replace-cross Abstract: Math word problems (MWPs) are critical K-12 educational tools, and customizing them to students' interests and ability levels can enhance lear

applicationsarxiv-cs-ai
14 Apr 2026
Safety

Endogenous Information in Routing Games: Memory-Constrained Equilibria, Recall Braess Paradoxes, and Memory Design

DGX agent

arXiv:2604.11733v1 Announce Type: cross Abstract: We study routing games in which travelers optimize over routes that are remembered or surfaced, rather than over a fixed exogenous action set. The pap

safetyarxiv-cs-ai
14 Apr 2026
Tutorials

Engineering Resource-constrained Software Systems with DNN Components: a Concept-based Pruning Approach

DGX agent

arXiv:2604.09988v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) are widely used by engineers to solve difficult problems that require predictive modeling from data. However, these models

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

EviRCOD: Evidence-Guided Probabilistic Decoding for Referring Camouflaged Object Detection

DGX agent

arXiv:2604.10894v1 Announce Type: new Abstract: Referring Camouflaged Object Detection (Ref-COD) focuses on segmenting specific camouflaged targets in a query image using category-aligned references.

model-releasesarxiv-cs-cv
14 Apr 2026
Research

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE

DGX agent

arXiv:2603.06003v2 Announce Type: replace Abstract: Sparse Mixture-of-Experts (SMoE) language models achieve strong capability at low per-token compute, yet deployment remains constrained by memory fo

researcharxiv-cs-lg
14 Apr 2026
Model Releases

Exploring Cross-Modal Flows for Few-Shot Learning

DGX agent

arXiv:2510.14543v4 Announce Type: replace Abstract: Aligning features from different modalities, is one of the most fundamental challenges for cross-modal tasks. Although pre-trained vision-language m

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Exploring the best way for UAV visual localization under Low-altitude Multi-view Observation Condition: a Benchmark

DGX agent

arXiv:2503.10692v2 Announce Type: replace Abstract: Absolute Visual Localization (AVL) enables an Unmanned Aerial Vehicle (UAV) to determine its position in GNSS-denied environments by establishing ge

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…636637638639640…1371
Next →