채널 상세

AI Papers Explained

40
구독자
38
영상
최근 영상
LONGER: Scaling Ultra-Long Sequence Modeling for Industrial Recommenders at ByteDance
LONGER: Scaling Ultra-Long Sequence Modeling for Industrial Recommenders at ByteDance
조회 17 좋아요 2 4개월 전
AIRA-Compose and AIRA-Design: LLM Agents Discover Novel Neural Architectures Beyond Transformer
AIRA-Compose and AIRA-Design: LLM Agents Discover Novel Neural Architectures Beyond Transformer
조회 3 좋아요 0 4개월 전
LIFE: A Unified Survey of Collaboration, Failure Attribution, and Self-Evolution in LLM Agents
LIFE: A Unified Survey of Collaboration, Failure Attribution, and Self-Evolution in LLM Agents
조회 6 좋아요 1 4개월 전
SilverTorch: Unified GPU Model-Based Serving for Large-Scale Recommendation at Meta
SilverTorch: Unified GPU Model-Based Serving for Large-Scale Recommendation at Meta
조회 4 좋아요 0 4개월 전
Explicit n^1.014 Lower Bound for the Erdős Unit Distance Problem via Golod-Shafarevich
Explicit n^1.014 Lower Bound for the Erdős Unit Distance Problem via Golod-Shafarevich
조회 7 좋아요 0 4개월 전
Code as Agent Harness: A Unified View of Executable, Verifiable, Stateful Agent Systems
Code as Agent Harness: A Unified View of Executable, Verifiable, Stateful Agent Systems
조회 1 좋아요 0 4개월 전
The RL Conductor: Training a 7B Model to Orchestrate LLM Agents via Reinforcement Learning
The RL Conductor: Training a 7B Model to Orchestrate LLM Agents via Reinforcement Learning
조회 7 좋아요 1 4개월 전
HeavySkill: Internalizing Parallel Reasoning and Summarization as an Inner LLM Skill
HeavySkill: Internalizing Parallel Reasoning and Summarization as an Inner LLM Skill
조회 3 좋아요 0 4개월 전
PLUM: Adapting Pre-trained LLMs for YouTube-Scale Generative Recommendations
PLUM: Adapting Pre-trained LLMs for YouTube-Scale Generative Recommendations
조회 15 좋아요 1 4개월 전
OneRec: Unifying Retrieval and Ranking with a Generative Recommender and DPO Alignment
OneRec: Unifying Retrieval and Ranking with a Generative Recommender and DPO Alignment
조회 8 좋아요 0 4개월 전
HSTU: Trillion-Parameter Generative Recommenders That Beat DLRMs at Scale
HSTU: Trillion-Parameter Generative Recommenders That Beat DLRMs at Scale
조회 22 좋아요 1 4개월 전
Constitutional AI: Training Harmless Assistants with AI Feedback Instead of Human Labels
Constitutional AI: Training Harmless Assistants with AI Feedback Instead of Human Labels
조회 9 좋아요 1 4개월 전
PaLM: Scaling a 540B Parameter Language Model with Pathways
PaLM: Scaling a 540B Parameter Language Model with Pathways
조회 29 좋아요 0 4개월 전
Chinchilla: Training Compute-Optimal Large Language Models
Chinchilla: Training Compute-Optimal Large Language Models
조회 7 좋아요 0 4개월 전
Vision Transformer (ViT): Transformers for Image Recognition at Scale
Vision Transformer (ViT): Transformers for Image Recognition at Scale
조회 5 좋아요 0 4개월 전
Recurrent Neural Network Regularization: Applying Dropout to LSTMs
Recurrent Neural Network Regularization: Applying Dropout to LSTMs
조회 22 좋아요 0 4개월 전
Generative Adversarial Nets: Goodfellow et al.'s Original GAN Paper Explained
Generative Adversarial Nets: Goodfellow et al.'s Original GAN Paper Explained
조회 0 좋아요 0 4개월 전
Latent Diffusion Models: High-Resolution Image Synthesis in Compressed Latent Space
Latent Diffusion Models: High-Resolution Image Synthesis in Compressed Latent Space
조회 5 좋아요 0 4개월 전
Batch Normalization: Reducing Internal Covariate Shift to Accelerate Deep Network Training
Batch Normalization: Reducing Internal Covariate Shift to Accelerate Deep Network Training
조회 1 좋아요 0 4개월 전
Neural Turing Machines: Differentiable Memory for Learning Algorithms
Neural Turing Machines: Differentiable Memory for Learning Algorithms
조회 10 좋아요 0 4개월 전
Playing Atari with Deep Reinforcement Learning: The Original DQN Paper
Playing Atari with Deep Reinforcement Learning: The Original DQN Paper
조회 5 좋아요 0 4개월 전
Auto-Encoding Variational Bayes: The Original VAE Paper by Kingma and Welling
Auto-Encoding Variational Bayes: The Original VAE Paper by Kingma and Welling
조회 27 좋아요 0 4개월 전
Bahdanau et al. (2014): Neural Machine Translation by Jointly Learning to Align and Translate
Bahdanau et al. (2014): Neural Machine Translation by Jointly Learning to Align and Translate
조회 4 좋아요 0 4개월 전
Adam: A Method for Stochastic Optimization (Kingma & Ba, 2015)
Adam: A Method for Stochastic Optimization (Kingma & Ba, 2015)
조회 10 좋아요 2 4개월 전
Scaling Laws for Neural Language Models: Power-Law Trends in Loss, Size, Data, and Compute
Scaling Laws for Neural Language Models: Power-Law Trends in Loss, Size, Data, and Compute
조회 11 좋아요 0 4개월 전
LoRA: Low-Rank Adaptation for Efficient Fine-Tuning of Large Language Models
LoRA: Low-Rank Adaptation for Efficient Fine-Tuning of Large Language Models
조회 17 좋아요 0 4개월 전
The Coffee Automaton: Quantifying the Rise and Fall of Complexity in Closed Systems
The Coffee Automaton: Quantifying the Rise and Fall of Complexity in Closed Systems
조회 1 좋아요 0 4개월 전
Deep Residual Learning for Image Recognition: The ResNet Paper Explained
Deep Residual Learning for Image Recognition: The ResNet Paper Explained
조회 15 좋아요 2 4개월 전
Order Matters: Extending Seq2Seq to Handle Sets as Inputs and Outputs
Order Matters: Extending Seq2Seq to Handle Sets as Inputs and Outputs
조회 3 좋아요 0 4개월 전
Denoising Diffusion Probabilistic Models (DDPM): High-Quality Image Synthesis Explained
Denoising Diffusion Probabilistic Models (DDPM): High-Quality Image Synthesis Explained
조회 4 좋아요 0 4개월 전
보기 방식
채널검색