爱可可-爱生活
24-09-21 22:12 微博认证:AI博主 2025微博新锐新知博主

几篇论文实现代码:
《VPTQ: Extreme Low-bit Vector Post-Training Quantization for Large Language Models》(EMNLP 2024) GitHub: github.com/microsoft/VPTQ
《AMEGO: Active Memory from long EGOcentric videos》(ECCV 2024) GitHub: github.com/gabrielegoletto/AMEGO [fig4]
《MotionFix: Text-Driven 3D Human Motion Editing》(SIGGRAPH Asia 2024) GitHub: github.com/athn-nik/motionfix
《3DTopia-XL: High-Quality 3D PBR Asset Generation via Primitive Diffusion》(2024) GitHub: github.com/3DTopia/3DTopia-XL
《Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think》(2024) GitHub: github.com/VisualComputingInstitute/diffusion-e2e-ft [fig1]
《GRIN: GRadient-INformed MoE》(2024) GitHub: github.com/microsoft/GRIN-MoE
《Oryx MLLM: On-Demand Spatial-Temporal Understanding at Arbitrary Resolution》(2024) GitHub: github.com/Oryx-mllm/Oryx [fig2]
《A Controlled Study on Long Context Extension and Generalization in LLMs》(2024) GitHub: github.com/Leooyii/LCEG [fig3]
《MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines》(2024) GitHub: github.com/CaraJ7/MMSearch
《LightningDrag: Lightning Fast and Accurate Drag-based Image Editing Emerging from Videos》(2024) GitHub: github.com/magic-research/LightningDrag
《Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning》(2024) GitHub: github.com/ytyz1307zzh/RefAug
《PDMX: A Large-Scale Public Domain MusicXML Dataset for Symbolic Music Processing》(2024) GitHub: github.com/pnlong/PDMX
《3D reconstruction with fast dipole sums》(2024) GitHub: github.com/cmu-ci-lab/fast_dipole_sums
《ABQ-LLM: Arbitrary-Bit Quantized Inference Acceleration for Large Language Models》(2024) GitHub: github.com/bytedance/ABQ-LLM
《Scaling Value Iteration Networks》(2024) GitHub: github.com/lucidrains/scaling-vin-pytorch [fig5]
《Effective Pre-Training of Audio Transformers for Sound Event Detection》(2024) GitHub: github.com/fschmid56/PretrainedSED
《MM3DGS SLAM: Multi-modal 3D Gaussian Splatting for SLAM Using Vision, Depth, and Inertial Measurements》(2024) GitHub: github.com/VITA-Group/MM3DGS-SLAM [fig6]
《Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models》(2024) GitHub: github.com/orionw/promptriever
《Vec2Face: Scaling Face Dataset Generation with Loosely Constrained Vectors》(2024) GitHub: github.com/HaiyuWu/Vec2Face
《DSBench: How Far Are Data Science Agents to Becoming Data Science Experts?》(2024) GitHub: github.com/LiqiangJing/DSBench
《Measuring audio prompt adherence with distribution-based embedding distances》(2024) GitHub: github.com/SonyCSLParis/audio-metrics
《On the Diagram of Thought》(2024) GitHub: github.com/diagram-of-thought/diagram-of-thought
《OPUS: Occupancy Prediction Using a Saprse Set》(2024) GitHub: github.com/jbwang1997/OPUS [fig7]

发布于 北京