AndroidLab
AndroidLab: First ever systematic benchmark for Android mobile agents shows that small, fine-tuned open models can power a JARVIS system on your smartphone 📱🔥
I'm Aymeric ("m-ric"), Machine Learning Engineer at Hugging Face.
AndroidLab: First ever systematic benchmark for Android mobile agents shows that small, fine-tuned open models can power a JARVIS system on your smartphone 📱🔥
Anthropic just released a chunk improvement technique that vastly improves RAG performance! 🔥
🌟 Cohere releases Aya 8B & 32B: SOTA multilingual models for 23 languages.
Categories: 🤖 Agents Complete: Terminé Key insights: Claude-3.5-Sonnet is really strong on agentic behaviour, much stronger than GPT-4o. Publication: 2024/10/22 Rating: ⭐️⭐️ Read: 2024/10/22
🧠 CLEAR: first multimodal benchmark to make models forget what we want them to forget
How to re-rank your snippets in RAG ⇒ ColBERT vs Rerankers vs Cross-Encoders ⚔️
image.png
Transformers v4.45.0 released: includes a lightning-fast method to build tools! ⚡️
🤗 I’m very proud to have supported CGIAR and Digital Green in making Farmer.chat, an app that supports 20k smallholder farmers on a daily basis 🌾
🚀 Hunyuan-Large just released by Tencent: Largest ever open MoE LLM, only 52B active parameters but beats LLaMA 3.1-405B on most academic benchmarks