AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
22 Apr 2026

Qwen3.6-27B-TQ3_4S is insanely good! https://huggingface.co/YTan2000/Qwen3.6-27B-TQ3_4S fit on my 16GB with 32k context Two prompts and I ge…

IndustryDGX agent

Qwen3.6-27B-TQ3_4S is a quantized 27 billion parameter language model that fits on 16GB of VRAM while supporting a 32k token context window, demonstrating strong performance across tested prompts. The

ReefNet: A Large-Scale Dataset and Benchmark for Fine-Grained Coral Reef Recognition

Model ReleasesDGX agent

arXiv:2510.16822v3 Announce Type: replace-cross Abstract: Coral reefs are rapidly declining under anthropogenic pressures (e.g., climate change), creating an urgent need for scalable and automated mon

Seeing Candidates at Scale: Multimodal LLMs for Visual Political Communication on Instagram

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.19489v1 Announce Type: new Abstract: This paper presents a computational case study that evaluates the capabilities of specialized machine learning models and emerging multimodal large lang

ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2604.19254v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) reduces the training cost of full-parameter fine-tuning for large language models (LLMs) by training only a sma

SpecAgent: A Speculative Retrieval and Forecasting Agent for Code Completion

Model ReleasesDGX agent

arXiv:2510.17925v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at code-related tasks but often struggle in realistic software repositories, where project-specific APIs an

Step 3.5 Flash is now live for Nous Portal users, free for the next 10 days. If you're running Hermes Agent with Nous Portal as your provide…

AgentsDGX agent

Step 3.5 Flash is now live for Nous Portal users, free for the next 10 days. If you're running Hermes Agent with Nous Portal as your provider, run 'hermes update' and then 'hermes model' to configure

StochasTok: Improving Fine-Grained Subword Understanding in LLMs

ResearchDGX agent

arXiv:2506.01687v3 Announce Type: replace Abstract: Subword-level understanding is integral to numerous tasks, including understanding multi-digit numbers, spelling mistakes, abbreviations, rhyming, a

Structure-Semantic Decoupled Modulation of Global Geospatial Embeddings for High-Resolution Remote Sensing Mapping

ResearchDGX agent

arXiv:2604.19591v1 Announce Type: new Abstract: Fine-grained high-resolution remote sensing mapping typically relies on localized visual features, which restricts cross-domain generalizability and oft

Symbolic Quantile Regression for the Interpretable Prediction of Conditional Quantiles

SafetyDGX agent

arXiv:2508.08080v2 Announce Type: replace Abstract: Symbolic Regression (SR) is a well-established framework for generating interpretable or white-box predictive models. Although SR has been successfu

Time Series Augmented Generation for Financial Applications

Model ReleasesDGX agent

arXiv:2604.19633v1 Announce Type: new Abstract: Evaluating the reasoning capabilities of Large Language Models (LLMs) for complex, quantitative financial tasks is a critical and unsolved challenge. St

TRN-R1-Zero: Text-rich Network Reasoning via LLMs with Reinforcement Learning Only

SafetyDGX agent

arXiv:2604.19070v1 Announce Type: new Abstract: Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure wi

Try Kimi K2.6 now on the AI Native Cloud: http://www.together.ai/models/kimi-k26#

ToolsDGX agent

Kimi K2.6 is now available for use on Together AI's cloud platform, which offers AI model deployment and inference services. This announcement indicates the model has been added to Together AI's roste

Tstars-Tryon 1.0: Robust and Realistic Virtual Try-On for Diverse Fashion Items

Model ReleasesDGX agent

arXiv:2604.19748v1 Announce Type: new Abstract: Recent advances in image generation and editing have opened new opportunities for virtual try-on. However, existing methods still struggle to meet compl

Two-dimensional early exit optimisation of LLM inference

Model ReleasesDGX agent

arXiv:2604.18592v1 Announce Type: cross Abstract: We introduce a two-dimensional (2D) early exit strategy that coordinates layer-wise and sentence-wise exiting for classification tasks in large langua

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search

Local AiDGX agent

arXiv:2604.19440v1 Announce Type: new Abstract: Recent work has demonstrated the promise of orchestrating large language models (LLMs) within evolutionary and agentic optimization systems. However, th

What’s new with compute: Scaling core and agentic workloads

Model ReleasesDGX agent

At Google Cloud Next, we’re announcing a range of compute capabilities to enable your core general purpose and AI workloads for the agentic world with higher performance and lower costs. Why it matter

When Does Verification Pay Off? A Closer Look at LLMs as Solution Verifiers

ResearchDGX agent

arXiv:2512.02304v2 Announce Type: replace Abstract: Large language models (LLMs) can act as both problem solvers and solution verifiers, where the latter select high-quality answers from a pool of sol

21 Apr 2026

Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization

Model ReleasesDGX agent

arXiv:2602.20743v2 Announce Type: replace Abstract: Anonymizing textual documents is a highly context-sensitive problem: the appropriate balance between privacy protection and utility preservation var

Agent-World: Scaling Real-World Environment Synthesis for Evolving General Agent Intelligence

AgentsDGX agent

arXiv:2604.18292v1 Announce Type: cross Abstract: Large language models are increasingly expected to serve as general-purpose agents that interact with external, stateful tool environments. The Model

Agentic Risk-Aware Set-Based Engineering Design

Model ReleasesDGX agent

arXiv:2604.16687v1 Announce Type: cross Abstract: This paper introduces a multi-agent framework guided by Large Language Models (LLMs) to assist in the early stages of engineering design, a phase ofte

An Integrated Deep-Learning Framework for Peptide-Protein Interaction Prediction and Target-Conditioned Peptide Generation with ConGA-PePPI and TC-PepGen

Model ReleasesDGX agent

arXiv:2604.18467v1 Announce Type: new Abstract: Motivation: Peptide-protein interactions (PepPIs) are central to cellular regulation and peptide therapeutics, but experimental characterization remains

An `Inverse' Experimental Framework to Estimate Market Efficiency

SafetyDGX agent

arXiv:2604.18130v1 Announce Type: new Abstract: Digital marketplaces processing billions of dollars annually represent critical infrastructure in sociotechnical ecosystems, yet their performance optim

An Uncertainty-Aware Loss Function Incorporating Fuzzy Logic: Application to MRI Brain Image Segmentation

Model ReleasesDGX agent

arXiv:2604.16490v1 Announce Type: new Abstract: Accurate brain image segmentation, particularly for distinguishing various tissues from magnetic resonance imaging (MRI) images, plays a pivotal role in

Anthropic gets $5B investment from Amazon, will use it to buy Amazon chips

Model ReleasesDGX agent

Amazon announced a 5 billion investment in Anthropic, with up to 20 billion more tied to commercial milestones. Anthropic committed to spending over $100 billion on AWS technologies over the next deca

AntiPaSTO: Self-Supervised Honesty Steering via Anti-Parallel Representations

Model ReleasesDGX agent

arXiv:2601.07473v4 Announce Type: replace Abstract: As models grow more capable, humans cannot reliably verify what they say. Scalable steering requires methods that are internal, self-supervised, and

Are Emotion and Rhetoric Neurons in LLM? Neuron Recognition and Adaptive Masking for Emotion-Rhetoric Prediction Steering

ResearchDGX agent

arXiv:2604.17255v1 Announce Type: new Abstract: Accurate comprehension and controllable generation of emotion and rhetoric are pivotal for enhancing the reasoning capabilities of large language models

Attention Is not Everything: Efficient Alternatives for Vision

ResearchDGX agent

arXiv:2604.17439v1 Announce Type: new Abstract: Recently computer vision has seen advancements mainly thanks to Transformer-based models. However many non-Transformer methods are still doing well bein

AWPD: Frequency Shield Network for Agnostic Watermark Presence Detection

Model ReleasesDGX agent

arXiv:2603.06723v3 Announce Type: replace Abstract: Invisible watermarks, as an essential technology for image copyright protection, have been widely deployed with the rapid development of social medi

Can LLM-Generated Text Empower Surgical Vision-Language Pre-training?

Model ReleasesDGX agent

arXiv:2604.18134v1 Announce Type: new Abstract: Recent advancements in self-supervised learning have led to powerful surgical vision encoders capable of spatiotemporal understanding. However, extendin

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction

Model ReleasesDGX agent

arXiv:2512.11988v3 Announce Type: replace Abstract: Accurate capture of human-object interaction from ubiquitous sensors like RGB cameras is important for applications in human understanding, gaming,

Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction

Model ReleasesDGX agent

arXiv:2604.16886v1 Announce Type: new Abstract: Generalist embodied agents must perform interactive, causally-dependent reasoning, continually interacting with the environment, acquiring information,

ComPASS: Towards Personalized Agentic Social Support via Tool-Augmented Companionship

Model ReleasesDGX agent

arXiv:2604.18356v1 Announce Type: new Abstract: Developing compassionate interactive systems requires agents to not only understand user emotions but also provide diverse, substantive support. While r

Conformal Prediction-Based MPC for Stochastic Linear Systems

ResearchDGX agent

arXiv:2512.10738v2 Announce Type: replace-cross Abstract: We propose a stochastic model predictive control (MPC) framework for linear systems subject to joint-in-time chance constraints under unknown

ConforNets: Latents-Based Conformational Control in OpenFold3

ResearchDGX agent

arXiv:2604.18559v1 Announce Type: cross Abstract: Models from the AlphaFold (AF) family reliably predict one dominant conformation for most well-ordered proteins but struggle to capture biologically r

Continual Safety Alignment via Gradient-Based Sample Selection

SafetyDGX agent

arXiv:2604.17215v1 Announce Type: new Abstract: Large language models require continuous adaptation to new tasks while preserving safety alignment. However, fine-tuning on even benign data often compr

Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR

ResearchDGX agent

arXiv:2603.07084v2 Announce Type: replace-cross Abstract: Reward hacking is a form of misalignment in which models overoptimize proxy rewards without genuinely solving the underlying task. Precisely m

Cross-Modal Bayesian Low-Rank Adaptation for Uncertainty-Aware Multimodal Learning

Model ReleasesDGX agent

arXiv:2604.16657v1 Announce Type: new Abstract: Large pre-trained language models are increasingly adapted to downstream tasks using parameter-efficient fine-tuning (PEFT), but existing PEFT methods a

Decomposing the Depth Profile of Fine-Tuning

ResearchDGX agent

arXiv:2604.17177v1 Announce Type: new Abstract: Fine-tuning adapts pretrained networks to new objectives. Whether the resulting depth profile of representational change reflects an intrinsic property

DEM Refinement and Validation on the Lunar Surface Using Shape-from-Shading with Chandrayaan-2 OHRC Imagery

Model ReleasesDGX agent

arXiv:2604.17436v1 Announce Type: new Abstract: This study presents a Shape from Shading (SfS) framework to enhance sub-metre resolution lunar digital elevation models (DEMs) using imagery from the Or

Depth Registers Unlock W4A4 on SwiGLU: A Reader/Generator Decomposition

Model ReleasesDGX agent

arXiv:2604.18128v1 Announce Type: new Abstract: We study post-training W4A4 quantization in a controlled 300M-parameter SwiGLU decoder-only language model trained on 5B tokens of FineWeb-Edu, and ask

Diffusion-Based Optimization for Accelerated Convergence of Redundant Dual-Arm Minimum Time Problems

ResearchDGX agent

arXiv:2604.16670v1 Announce Type: new Abstract: We present a framework leveraging a novel variant of the model-based diffusion algorithm to minimize the time required for a redundant dual-arm robot co

Domain-oriented RAG Assessment (DoRA): Synthetic Benchmarking for RAG-based Question Answering on Defense Documents

Model ReleasesDGX agent

arXiv:2604.17943v1 Announce Type: new Abstract: Open-domain RAG benchmarks over public corpora can overestimate deployment performance due to pretraining overlap and weak attribution requirements. We

DREAM: Dynamic Retinal Enhancement with Adaptive Multi-modal Fusion for Expert Precision Medical Report Generation

Model ReleasesDGX agent

arXiv:2604.17209v1 Announce Type: new Abstract: Automating medical reports for retinal images requires a sophisticated blend of visual pattern recognition and deep clinical knowledge. Current Large Vi

Driving in Corner Case: A Real-World Adversarial Closed-Loop Evaluation Platform for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2512.16055v2 Announce Type: replace Abstract: Safety-critical corner cases, difficult to collect in the real world, are crucial for evaluating end-to-end autonomous driving. Adversarial interact

Dual-stream Spatio-Temporal GCN-Transformer Network for 3D Human Pose Estimation

Model ReleasesDGX agent

arXiv:2604.17688v1 Announce Type: new Abstract: 3D human pose estimation is a classic and important research direction in the field of computer vision. In recent years, Transformer-based methods have

DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies

ResearchDGX agent

arXiv:2503.14324v3 Announce Type: replace-cross Abstract: The differing representation spaces required for visual understanding and generation pose a challenge in unifying them within the autoregressi

DuConTE: Dual-Granularity Text Encoder with Topology-Constrained Attention for Text-attributed Graphs

Model ReleasesDGX agent

arXiv:2604.17411v1 Announce Type: new Abstract: Text-attributed graphs integrate semantic information of node texts with topological structure, offering significant value in various applications such

EasyVideoR1: Easier RL for Video Understanding

Model ReleasesDGX agent

arXiv:2604.16893v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) has demonstrated remarkable effectiveness in improving the reasoning capabilities of large languag

End-to-end Listen, Look, Speak and Act

Model ReleasesDGX agent

arXiv:2510.16756v2 Announce Type: replace-cross Abstract: Human interaction is inherently multimodal and full-duplex: we listen while watching, speak while acting, and fluidly adapt to turn-taking and

Enhancing Zero-shot Personalized Image Aesthetics Assessment with Profile-aware Multimodal LLM

ResearchDGX agent

arXiv:2604.17233v1 Announce Type: new Abstract: Personalized image aesthetics assessment (PIAA) aims to predict an individual user's subjective rating of an image, which requires modeling user-specifi

Error as Signal: Stiffness-Aware Diffusion Sampling via Embedded Runge-Kutta Guidance

Model ReleasesDGX agent

arXiv:2603.03692v2 Announce Type: replace Abstract: Classifier-Free Guidance (CFG) has established the foundation for guidance mechanisms in diffusion models, showing that well-designed guidance proxi

Explanation Bias is a Product: Revealing the Hidden Lexical and Position Preferences in Post-Hoc Feature Attribution

SafetyDGX agent

arXiv:2512.11108v3 Announce Type: replace Abstract: Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a

FairLogue: Evaluating Intersectional Fairness across Clinical Machine Learning Use Cases using the All of Us Research Program

SafetyDGX agent

arXiv:2604.16450v1 Announce Type: cross Abstract: Intersectional biases in healthcare data can produce compound disparities in clinical machine learning models, yet most fairness evaluations assess de

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

SafetyDGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

FlashFPS: Efficient Farthest Point Sampling for Large-Scale Point Clouds via Pruning and Caching

Model ReleasesDGX agent

arXiv:2604.17720v1 Announce Type: cross Abstract: Point-based Neural Networks (PNNs) have become a key approach for point cloud processing. However, a core operation in these models, Farthest Point Sa

FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation

ResearchDGX agent

arXiv:2603.09721v2 Announce Type: replace Abstract: High-fidelity video generation remains challenging for diffusion models due to the difficulty of modeling complex spatio-temporal dynamics efficient

Frequency-guided Multi-level Reasoning for Scene Graph Generation in Video

ResearchDGX agent

arXiv:2604.17298v1 Announce Type: new Abstract: Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understandin

From 2:4 to 8:16 sparsity patterns in LLMs for Outliers and Weights with Variance Correction

ResearchDGX agent

arXiv:2507.03052v2 Announce Type: replace Abstract: As large language models (LLMs) grow in size, efficient compression techniques like quantization and sparsification are critical. While quantization

GeoRC: A Benchmark for Geolocation Reasoning Chains

Model ReleasesDGX agent

arXiv:2601.21278v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are good at recognizing the global location of a photograph -- their geolocation prediction accuracy rivals the

Harness Engineering Without the Hype: 🦄 AI That Works #54 https://x.com/i/broadcasts/1DxLdvQrDwkxm

AgentsDGX agent

This episode of 'AI That Works' discusses practical approaches to harness engineering in AI systems, likely focusing on techniques for effectively prompting and controlling AI model behavior beyond ma

← Previous
1…426427428429430…1059
Next →