7 тысяч подписчиков
206 видео
SOLAR 10.7B: Scaling LLMs with Depth Up-Scaling
Backpack Language Models
LoRA: Low Rank Adaptation of Large Language Models
Video
#209
LAION-5B: Dataset for training image-text models
Apple's OpenELM models
Microsoft Phi-3
Introduction to LangChain
#206
OLMo
CircuitVQA: A VQA Dataset for Electrical Circuit Images
XTR: ConteXtualized Token Retriever
The Reversal Curse for LLMs
LaMP: Personalization Benchmark for LLMs
#205
CALM: LLM Augmented LLMs
BitNet: Scaling 1-bit Transformers for LLMs
DeciLM: 15x higher throughput than Llama 2
AudioGen: Textually Guided Audio Generation
Part-1 of our IJCAI tutorial on deep learning for Brain Encoding and Decoding
Codex: Large Language Models Trained on Code
Google Deepmind's Gemma
Google Flan: Finetuned Language Models are zero shot learners
Video-to-video translation, video object tracking, video super-resolution using CoDeF
OpenAI DALL·E 2: Hierarchical text conditional image generation with clip latents
#274
MEGALODON: Efficient LLM Pretraining and Inference with Unlimited Context Length
Microsoft VASA-1: Lifelike Audio Driven Talking Faces Generated in Real Time
KAN: Kolmogorov-Arnold Networks
Interpreting Stable Diffusion Using Cross Attention
Part-4 of our IJCAI tutorial on deep learning for Brain Encoding and Decoding
Are Emergent Abilities of Large Language Models a Mirage?
#217
Supporting Infinite Context Length using TransformerFAM
YOCO: Decoder-Decoder Architectures for LLMs
Imagen: Photorealistic Text to Image Diffusion Models with Deep Language Understanding
Mamba: Linear-Time Sequence Modeling with Selective State Spaces
OpenAI's gpt4o
OAK: Enriching Doc Representations using Auxiliary Knowledge for XC
PLLaVA: Parameter-free LLaVA Extension from Images to Videos for Video Dense Captioning
Infini attention and Infini Transformer
#204
Stable Diffusion High resolution image synthesis with latent diffusion models
DeltaLM: Encoder Decoder Pre training for Language Generation and Translation
LLMLingua: Compressing Prompts for Accelerated Inference of LLMs
OpenAI Whisper: Robust Speech Recognition via Large Scale Weak Supervision