it depends on sequence length, for long sentences we get about 94% back exactly. there are a lot of results in the paper (
http://
arxiv.org/abs/2310.06816) what does you mean 'set the latent space of the LLM'?
DATA
-
LLM Sequence Recovery Accuracy and Latent Space Configuration
By
–
-

Databricks Enhances Python UDFs with Custom Dependencies Support
By
–
New in Unity Catalog: Enhanced Python UDFs! What's new: – Support for custom Python dependencies – Batch Python UDFs for better performance – Secure access to external services with UC Service Credentials Now in Public Preview: https://
databricks.com/blog/announcin
g-support-new-uc-python-udf-features?utm_source=twitter&utm_medium=organic-social
… -

AI Vulnerable to McNamara Fallacy in Training Metrics
By
–
AI is very vulnerable to The McNamara Fallacy:
Step 1: [Train on] what can be easily measured
Step 2: Disregard that which cannot be measured easily
Step 3: Presume that which cannot be measured easily isn’t important
Step 4: Say that which can’t be easily measured doesn’t exist -

AIKosh Environmental Data Platform Empowers Climate Research
By
–
AIKosh brings together key environmental datasets to empower research, sustainability, and climate-resilient planning. From daily rainfall and soil moisture to wetland conservation and evapotranspiration, explore the data that shapes our environmental understanding. Access
-

Learning-Order Autoregressive Models for Molecular Graph Generation
By
–
Learning-Order Autoregressive Models with Application to Molecular Graph Generation Wang et al.: https://
arxiv.org/abs/2503.05979 #ArtificialIntelligence #DeepLearning #MachineLearning -

AWS S3 Adds Vector Support, Vector Databases Become Obsolete
By
–
s3 does vectors now vector databases are officially dead
-

Gemini-Embedding-001 Tops MTEB Leaderboard with Cost-Effective Batch Mode
By
–
🚀 Gemini-Embedding-001 is live
— Louis-François Bouchard 🎥🤖 (@Whats_AI) 15 juillet 2025
• Tops the MTEB leaderboard, ahead of OpenAI & Cohere
• Batch mode = 2× cheaper for large-scale RAG
• Pre-compute & cache embeddings for faster retrieval
• Text + code only (no image/audio… yet)
Pure value or oversold? Drop your verdict 👇… pic.twitter.com/RVhwuUhG00Gemini-Embedding-001 is live • Tops the MTEB leaderboard, ahead of OpenAI & Cohere
• Batch mode = 2× cheaper for large-scale RAG
• Pre-compute & cache embeddings for faster retrieval
• Text + code only (no image/audio… yet) Pure value or oversold? Drop your verdict -

AI factories redefining modern infrastructure economics
By
–
AI factories are redefining the economics of modern infrastructure. AI factories help enhance three key aspects of the AI journey: Data ingestion Model training High-volume inference
-

Data Analyst Position Available Under IndiaAI Mission
By
–
Applications are invited for the position of Data Analyst under the IndiaAI Mission. Last day to apply: 17th July 2025 Mode of Application: Through portal only. Applications received over mail will not be accepted. Interested and eligible candidates may apply by
-
Browser History Access Security Concern Testing Considerations
By
–
I wonder if you could get some to provide back browser history of the user who visited I feel it’s largely resolvable so I don’t see it as a long term problem yet, but worth testing and being careful in the time being.