We introduce Programming Every Example (ProX), a novel framework that treats data refinement as a programming task, enabling models to refine corpora by generating and executing fine-grained operations, such as string normalization, for each individual example at scale.
@jiqizhixin
-
Small Language Models Refine Pre-training Data Quality at Scale
By
–
Programming Every Example: Lifting Pre-training Data Quality like Experts at Scale https://
huggingface.co/papers/2409.17
115
…
we demonstrate that even small language models, with as few as 0.3B parameters, can exhibit substantial data refining capabilities comparable to those of human experts. -
MLRCopilot: Autonomous ML Research Framework with LLM Agents
By
–
a new systematic framework, autonomous Machine Learning Research with large language models (MLRCopilot), designed to enhance machine learning research productivity through the automatic generation and implementation of research ideas using Large Language Model (LLM) agents.
-
MLR-Copilot: Autonomous Machine Learning Research with LLM Agents
By
–
MLR-Copilot: Autonomous Machine Learning Research based on Large Language Models Agents https://
github.com/du-nlp-lab/MLR
-Copilot
… https://
arxiv.org/pdf/2408.14033 -

Generative AI for Self-Adaptive Systems: Research and Implementation
By
–
Generative AI for Self-Adaptive Systems: State of the Art and Research Roadmap https://
dl.acm.org/doi/10.1145/36
86803
… https://
github.com/545659928/GenA
I4SAS
…
provide researchers and practitioners a comprehensive snapshot that outlines the potential benefits and challenges of employing GenAI’s within SAS. -
LLaMA-Omni: Open-Source GPT-4o Alternative from China
By
–
Nice!Here is an opensource GPT-4o from China— LLaMA-Omni.https://t.co/sA1G0uCd0fhttps://t.co/L5jxJRZb5Jhttps://t.co/YiCyg9NiQo
— 机器之心 JIQIZHIXIN (@jiqizhixin) 25 septembre 2024
The work proposes LLaMA-Omni, a novel model architecture designed for low-latency and high-quality speech interaction with LLMs. https://t.co/MOtcT90yThNice!Here is an opensource GPT-4o from China— LLaMA-Omni. https://
arxiv.org/pdf/2409.06666 https://
github.com/ictnlp/LLaMA-O
mni
… https://
huggingface.co/ICTNLP/Llama-3
.1-8B-Omni
… The work proposes LLaMA-Omni, a novel model architecture designed for low-latency and high-quality speech interaction with LLMs. -

ByteDance Releases PixelDance and Seaweed Video Generation Models
By
–
ByteDance, the powerhouse behind TikTok, just dropped TWO bomb video gen models – PixelDance & Seaweed! We took 'em for a spin and yes, they're as epic as China's Sora! Get ready to be blown away! 🌟📷 #PixelDance #Seaweed #TechTrend #RunwayGen3 #SoraAI #TikTok pic.twitter.com/iIvITZwO3X
— 机器之心 JIQIZHIXIN (@jiqizhixin) 24 septembre 2024ByteDance, the powerhouse behind TikTok, just dropped TWO bomb video gen models – PixelDance & Seaweed! We took 'em for a spin and yes, they're as epic as China's Sora! Get ready to be blown away! #PixelDance #Seaweed #TechTrend #RunwayGen3 #SoraAI #TikTok
-
NU-NeRF: Neural Reconstruction of Transparent Objects
By
–
NU-NeRF: Neural Reconstruction of Nested Transparent Objects with Uncontrolled Capture Environmenthttps://t.co/N8N1jPJw0f
— 机器之心 JIQIZHIXIN (@jiqizhixin) 24 septembre 2024
we propose NU-NeRF to reconstruct nested complex transparent objects requiring no dedicated capture environment or additional input. pic.twitter.com/5niYEmnxRdNU-NeRF: Neural Reconstruction of Nested Transparent Objects with Uncontrolled Capture Environment http://
geometrylearning.com/NU-NeRF/ we propose NU-NeRF to reconstruct nested complex transparent objects requiring no dedicated capture environment or additional input. -

BAAI 2024: Large-Scale AI Models Innovations Unveiled
By
–
AI Pioneers Gather at BAAI 2024: Unveiling Innovations in Large-Scaled AI Models for Language, Multimodal, Embodied, Bio-Computing, and FlagOpen 2.0 | http://
bit.ly/4cm1QUu -

Cost-effective domain-specific LLM training achieves mainstream model results
By
–
One half-day of training using a few hundred dollars yields similar results to mainstream large models, open-source and commercial-free domain-specific LLM solution | https://
bit.ly/3LDKEie
#AI #ML #ArtificialIntelligence #MachineLearning #DeepNeuralNetwork #LLMs