Jul 17, 2023Notes on Giving Robots a Hand: Learning Generalizable Manipulation with Eye-in-Hand Human Video DemonstrationsFeral Machine
Jul 17, 2023Notes on Instruction Mining: High-Quality Instruction Data Selection for Large Language ModelsFeral Machine
Jul 17, 2023Notes on SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Task PlanningFeral Machine
Jul 16, 2023Notes on Patch n' Pack: NaViT, a Vision Transformer for any Aspect Ratio and ResolutionFeral Machine
Jul 16, 2023Notes on Stack More Layers Differently: High-Rank Training Through Low-Rank UpdatesFeral Machine
Jul 16, 2023Notes on T2I-CompBench: A Comprehensive Benchmark for Open-world Compositional Text-to-image GenerationFeral Machine
Jul 16, 2023Notes on Domain-Agnostic Tuning-Encoder for Fast Personalization of Text-To-Image ModelsFeral Machine
Jul 16, 2023Notes on Animate-A-Story: Storytelling with Retrieval-Augmented Video GenerationFeral Machine
Jul 16, 2023Notes on Distilling Large Language Models for Biomedical Knowledge Extraction: A Case Study on Adverse Drug EventsFeral Machine
Jul 16, 2023Notes on InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and GenerationFeral Machine
Jul 16, 2023Notes on In-context Autoencoder for Context Compression in a Large Language ModelFeral Machine
Jul 16, 2023Notes on HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image ModelsFeral Machine