AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
28 Apr 2026

Reading in the Dark: Low-light Scene Text Recognition

Model ReleasesDGX agent

arXiv:2604.23685v1 Announce Type: new Abstract: Accurate text recognition in low-light environments is essential for intelligent systems in applications ranging from autonomous vehicles to smart surve

Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs

Model ReleasesDGX agent

arXiv:2509.25414v2 Announce Type: replace-cross Abstract: Large language models are often adapted using parameter-efficient techniques such as Low-Rank Adaptation (LoRA), formulated as y = W_0x + BAx,

Scalable LLM-based Coding of Dialogue in Healthcare Simulation: Balancing Coding Performance, Processing Time, and Environmental Impact

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.23255v1 Announce Type: cross Abstract: Research shows that dialogue, the interactive process through which participants articulate their thinking, plays a central role in constructing share

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

Model ReleasesDGX agent

arXiv:2604.23274v1 Announce Type: new Abstract: Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabe

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

Model ReleasesDGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

StereoFoley: Object-Aware Stereo Audio Generation from Video

ResearchDGX agent

We present StereoFoley, a video-to-audio generation framework that produces semantically aligned, temporally synchronized, and spatially accurate stereo sound at 48 kHz. While recent generative video-

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

Model ReleasesDGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

Model ReleasesDGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

Symbolic recovery of PDEs from measurement data

ResearchDGX agent

arXiv:2602.15603v2 Announce Type: replace Abstract: Models based on partial differential equations (PDEs) are powerful for describing a wide range of complex phenomena in the natural sciences. Accurat

Task-guided Spatiotemporal Network with Diffusion Augmentation for EEG-based Dementia Diagnosis and MMSE Prediction

ResearchDGX agent

arXiv:2604.23964v1 Announce Type: cross Abstract: Patients with dementia typically exhibit cognitive impairment, which is routinely assessed using the Mini-Mental State Examination (MMSE). Concurrentl

The Consensus Trap: Dissecting Subjectivity and the 'Ground Truth' Illusion in Data Annotation

SafetyDGX agent

arXiv:2602.11318v3 Announce Type: replace Abstract: In machine learning, 'ground truth' refers to the assumed correct labels used to train and evaluate models. However, the foundational 'ground truth'

Think Anywhere in Code Generation

AgentsDGX agent

arXiv:2603.29957v3 Announce Type: replace-cross Abstract: Recent advances in reasoning Large Language Models (LLMs) have primarily relied on upfront thinking, where reasoning occurs before final answe

Unrealized Expectations: Comparing AI Methods vs Classical Algorithms for Maximum Independent Set

Local AiDGX agent

arXiv:2502.03669v3 Announce Type: replace-cross Abstract: AI methods, such as generative models and reinforcement learning, have recently been applied to combinatorial optimization (CO) problems, espe

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

AgentsDGX agent

arXiv:2604.23674v1 Announce Type: new Abstract: With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people co

WebSerial Vision Training for Microcontrollers: A Browser-Based Companion to On-Device CNN Training

Model ReleasesDGX agent

arXiv:2604.22834v1 Announce Type: new Abstract: This paper presents webmcu-vision-web, a single-file, zero-install browser application for end-to-end TinyML vision model training and deployment on the

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

Model ReleasesDGX agent

arXiv:2604.23942v1 Announce Type: cross Abstract: People increasingly turn to large language models (LLMs) to interpret ambiguous social situations: a delayed text reply, an unusually cold supervisor,

When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL

Model ReleasesDGX agent

arXiv:2604.23398v1 Announce Type: new Abstract: We report a reproducible error pattern in GPT-5.4 on OWL~2~DL compliance queries: the model frequently answers ``unknown'' when the reasoner-entailed an

27 Apr 2026

Adapting MLLMs for Nuanced Video Retrieval

ResearchDGX agent

arXiv:2512.13511v2 Announce Type: replace Abstract: Our objective is to build an embedding model that captures the nuanced relationship between a search query and candidate videos. We cover three aspe

AgentBound: Securing Execution Boundaries of AI Agents

SafetyDGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography

Model ReleasesDGX agent

arXiv:2604.22139v1 Announce Type: new Abstract: Reliable automated analysis of Optical Coherence Tomography (OCT) imaging is crucial for diagnosing retinal disorders but faces a critical barrier: the

Chain-of-Memory: Lightweight Memory Construction with Dynamic Evolution for LLM Agents

ResearchDGX agent

arXiv:2601.14287v2 Announce Type: replace Abstract: External memory systems are pivotal for enabling Large Language Model (LLM) agents to maintain persistent knowledge and perform long-horizon decisio

CLVAE: A Variational Autoencoder for Long-Term Customer Revenue Forecasting

TutorialsDGX agent

arXiv:2604.22636v1 Announce Type: cross Abstract: Predicting customers' long-term revenue from sparse and irregular transaction data is central to marketing resource allocation in non-contractual sett

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

SafetyDGX agent

arXiv:2604.22542v1 Announce Type: cross Abstract: Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. T

Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation

ResearchDGX agent

arXiv:2510.19592v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) demonstrate strong video understanding by attending to visual tokens relevant to textual queries. To direct

EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges

Model ReleasesDGX agent

arXiv:2604.22595v1 Announce Type: new Abstract: CLIP has demonstrated strong generalization in visual domains through natural language supervision, even for video action recognition. However, most exi

@FireworksAI_HQ deeply believes in delivering the frontier quality. We will spend all the effort giving our users and the broad community th…

Model ReleasesDGX agent

@FireworksAI_HQ deeply believes in delivering the frontier quality. We will spend all the effort giving our users and the broad community the best OSS model quality. Deepseek V4 API up anytime now aft

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we exp…

AgentsDGX agent

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we explored what happens if we train an AI to do that job instead.

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

SafetyDGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

how to adjust the thinking effort for deepseek v4 on ollama cloud

Model ReleasesDGX agent

DeepSeek V4 models on Ollama Cloud support three thinking modes: 'No thinking' for fast answers, 'Thinking' for careful analysis, and 'Max thinking' for maximum reasoning effort . Users can adjust thi

LayerBoost: Layer-Aware Attention Reduction for Efficient LLMs

ResearchDGX agent

arXiv:2604.22050v1 Announce Type: cross Abstract: Transformers are mostly relying on softmax attention, which introduces quadratic complexity with respect to sequence length and remains a major bottle

LLMs as Assessors: Right for the Right Reason?

Model ReleasesDGX agent

arXiv:2601.08919v2 Announce Type: replace-cross Abstract: A good deal of recent research has focused on how Large Language Models (LLMs) may be used as judges in place of humans to evaluate the qualit

NiuTrans.LMT: Toward Inclusive and Scalable Multilingual Machine Translation with LLMs

ResearchDGX agent

arXiv:2511.07003v2 Announce Type: replace Abstract: Large language models have significantly advanced Multilingual Machine Translation (MMT), yet scaling to many languages while keeping quality robust

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

Model ReleasesDGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

PL-MTEB: Polish Massive Text Embedding Benchmark

Model ReleasesDGX agent

arXiv:2405.10138v2 Announce Type: replace Abstract: In this paper, we introduce the Polish Massive Text Embedding Benchmark (PL-MTEB), a comprehensive benchmark for text embeddings in the Polish langu

PreMoE: Proactive Inference for Efficient Mixture-of-Experts

ResearchDGX agent

arXiv:2505.17639v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models offer dynamic computation, but are typically deployed as static full-capacity models, missing opportunities for depl

PrivUn: Unveiling Latent Ripple Effects and Shallow Forgetting in Privacy Unlearning

ResearchDGX agent

arXiv:2604.22076v1 Announce Type: cross Abstract: Large language models (LLMs) often memorize private information during training, raising serious privacy concerns. While machine unlearning has emerge

Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition

Model ReleasesDGX agent

arXiv:2604.22390v1 Announce Type: new Abstract: Visual Place Recognition (VPR) determines a query image's geographic location by matching it against geotagged databases. However, existing methods stru

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

Model ReleasesDGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

Rethinking Math Reasoning Evaluation: A Robust LLM-as-a-Judge Framework Beyond Symbolic Rigidity

ResearchDGX agent

arXiv:2604.22597v1 Announce Type: new Abstract: Recent advancements in large language models have led to significant improvements across various tasks, including mathematical reasoning, which is used

Segment Any-Quality Images with Generative Latent Space Enhancement

ApplicationsDGX agent

arXiv:2503.12507v3 Announce Type: replace Abstract: Despite their success, Segment Anything Models (SAMs) experience significant performance drops on severely degraded, low-quality images, limiting th

SSG: Logit-Balanced Vocabulary Partitioning for LLM Watermarking

ResearchDGX agent

arXiv:2604.22438v1 Announce Type: cross Abstract: Watermarking has emerged as a promising technique for tracing the authorship of content generated by large language models (LLMs). Among existing appr

This is how it's done! Who else should we ask to release weights?

IndustryDGX agent

Clem Delangue, CEO of Hugging Face, discusses best practices for releasing model weights in the open-source AI community and advocates for other AI organizations to follow suit in making their models

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem

SafetyDGX agent

arXiv:2506.17299v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become increasingly deployed in safety-critical applications, the lack of systematic methods to assess their v

Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations

ResearchDGX agent

arXiv:2601.03779v2 Announce Type: replace Abstract: We explore intrinsic dimension (ID) of LLM representations as a marker of linguistic complexity. Specifically, we test whether ID differences across

UR^2: Unify RAG and Reasoning through Reinforcement Learning

Model ReleasesDGX agent

arXiv:2508.06165v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown strong capabilities through two complementary paradigms: Retrieval-Augmented Generation (RAG) for know

26 Apr 2026

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threa…

Model ReleasesDGX agent

7) Multi-agent design. I loved the design of Cove (they were the acquired by Microsoft). It was more of a whiteboard than tabs or chat threads. I don’t think we’ve cracked the right UI for managing ag

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate'…

Model ReleasesDGX agent

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST t…

Model ReleasesDGX agent

DeepSeek v4 Flash with *local inference* after 24h of playing with that: even with the 2 bit selective quantization GGUF, iti is the FIRST time I feel I have a frontier model running on my computer. T

25 Apr 2026

5.5 is so earnest 'little engine that could' energy

IndustryDGX agent

Sam Altman praised OpenAI's o1 model (version 5.5) for its earnest, persistent approach to problem-solving, comparing it to the 'little engine that could' mentality. The comment reflects Altman's pers

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹C…

Model ReleasesDGX agent

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹Claude Code: Set model to deepseek-v4-pro[1m] to unlock 1M conte

http://reddit.com/r/LocalLLaMA

ResearchDGX agent

r/LocalLLaMA is a subreddit community dedicated to discussing and sharing resources about running large language models locally on personal computers, covering topics like model optimization, hardware

WHY ARE YOU LIKE THIS

Model ReleasesDGX agent

@scottjla on Twitter in reply to my pelican riding a bicycle benchmark: I feel like we need to stack these tests now I checked to confirm that the model (ChatGPT Images 2.0) added the 'WHY ARE YOU LIK

24 Apr 2026

4⃣4⃣4⃣4⃣

Model ReleasesDGX agent

4⃣4⃣4⃣4⃣ Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit

A-THENA: Early Intrusion Detection for IoT with Time-Aware Hybrid Encoding and Network-Specific Augmentation

Model ReleasesDGX agent

arXiv:2604.21623v1 Announce Type: cross Abstract: The proliferation of Internet of Things (IoT) devices has significantly expanded attack surfaces, making IoT ecosystems particularly susceptible to so

Absorber LLM: Harnessing Causal Synchronization for Test-Time Training

Model ReleasesDGX agent

arXiv:2604.20915v1 Announce Type: cross Abstract: Transformers suffer from a high computational cost that grows with sequence length for self-attention, making inference in long streams prohibited by

AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security

Model ReleasesDGX agent

arXiv:2601.18491v2 Announce Type: replace Abstract: The rise of AI agents introduces complex safety and security challenges arising from autonomous tool use and environmental interactions. Current gua

Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents

Model ReleasesDGX agent

arXiv:2604.21134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently misread values, hallucinate details, and confuse overlapping elements in charts. Current approaches rely solely

BioMiner: A Multi-modal System for Automated Mining of Protein-Ligand Bioactivity Data from Literature

Model ReleasesDGX agent

arXiv:2604.21508v1 Announce Type: new Abstract: Protein-ligand bioactivity data published in the literature are essential for drug discovery, yet manual curation struggles to keep pace with rapidly gr

BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion

SafetyDGX agent

arXiv:2604.04395v2 Announce Type: replace Abstract: 3D conducting motion generation aims to synthesize fine-grained conductor motions from music, with broad potential in music education, virtual perfo

CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

Model ReleasesDGX agent

arXiv:2604.21308v1 Announce Type: cross Abstract: Enterprise LLM agents can dramatically improve workplace productivity, but their core capability, retrieving and using internal context to act on a us

← Previous
1…423424425426427…1059
Next →