/AI Weekly/Issue 67

Issue #67 18 stories

Global AI Weekly

Apple Intelligence is coming & Microsoft's Phi 3.5

Published Tuesday, August 27, 2024

In this issue

Highlights

4 stories
In this issue

Research

1 story
arxiv.org

Exploring mixture of experts models

As foundational models get bigger it becomes harder to train them. We simply don't have enough compute power to train models in a reasonable time and budget. To fix this, people have turned to mixture of experts models. But how do they work? This paper explains one approach to use mixture of experts. It's advanced, but cool to see how much power you can push out of a deep learning model!

In this issue

Video

4 stories
Designing High-Quality Synthetic Data for Training & Fine-tuning LLMs
youtube.com

Designing High-Quality Synthetic Data for Training & Fine-tuning LLMs

At this critical juncture in AI development, we face a scarcity of novel, high-quality data for training and fine-tuning large language models (LLMs). This shortage poses significant challenges for organizations aiming to enhance model performance or adapt them to specialized domains. In this hour-long livestream, Yev Meyer, Ph.D., Chief Scientist at Gretel, will showcase the company's innovative solution to this bottleneck.

In this issue

Articles

8 stories
In this issue

Code

1 story
Every Tuesday

Get the next issue in your inbox.

The AI links worth your time, curated by the community and always free.