For red teaming, the surface area of potential risks is extremely large. As a result, our approach involves extensive threat modeling, & automated evals to identify risks. Then our process enables capable experts to identify key risks, and report them back to model developers.
AI
-

Hybrid Approach to AI Capability Testing and Risk Evaluation
By
–
Our platform is built to both continually monitor and evaluate capabilities, and measure risks and vulnerabilities via red teaming. In particular, we believe in a hybrid approach to test and evaluation which involves both automated evaluation and expert evaluations.
-

Scale AI Launches LLM Test and Evaluation Platform
By
–
Today, alongside our collaboration with the @WhiteHouse and @DEFCON in an evaluation of the leading LLMs— @scale_ai is announcing the release of our Test and Evaluation platform and approach to enable for safe and scalable deployment of AI systems. Read thread for more
-
DEFCON Red Teams Leading AI Models for Security Vulnerabilities
By
–
At @DEFCON
, cybersecurity experts will be red teaming leading models from OpenAI, Anthropic, Google, and more. They will be using @scale_ai
’s platform to proactively identify vulnerabilities and report them to the model developers. -

Preparing for the Future of Work: Leadership and Culture
By
–
Preparing for the #FutureofWork https://
bit.ly/3WSqohk via @mercer #leadership #CompanyCulture -

Tidier: AI and Machine Learning Resource Link
By
–
Tidier https://
bit.ly/3DOG5NK
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

Drive Performance Through Data-Driven Collaboration and Innovation
By
–
See real-world customer stories that show how you can drive performance with a culture of collaboration, sharing, and innovation with data. Watch #AlteryxInspire sessions with leaders like @BankofAmerica
, @WestRock
, and @RoyalCaribbean today: https://
ow.ly/G1Zc50Pux5O -
Open Source AI Models: Predictability as Key Advantage
By
–
But it's adding a new layer to the discussion around the benefits of open source models. In addition to the usual stuff—latency, cost, privacy—they're increasingly attractive because they're just predictable.
-
API Dependencies: Trading Control for Convenience and Speed
By
–
That's, of course, the bargain you make when building a product on top of any API: you trade convenience and agility for control of the technology under your products. But if history is any indicator, a whole universe of companies are willing to make that trade.
-
OpenAI Faces Hardware Constraints Limiting GPT-4 Access
By
–
OpenAI is, like many other companies, facing hardware constraints. It's directionally honest about it. It limits the number of GPT-4 calls for Chat users. Developers have to pay for its APIs (like davinci or GPT 3.5-Turbo) for at least a month before they get access to GPT-4.
