Open Source AI Is the Path Forward

文摘 2024-07-24 13:11 广东

July 23, 2024

By Mark Zuckerberg, Founder and CEO

In the early days of high-performance computing, the major tech companies of the day each invested heavily in developing their own closed source versions of Unix. It was hard to imagine at the time that any other approach could develop such advanced software. Eventually though, open source Linux gained popularity – initially because it allowed developers to modify its code however they wanted and was more affordable, and over time because it became more advanced, more secure, and had a broader ecosystem supporting more capabilities than any closed Unix. Today, Linux is the industry standard foundation for both cloud computing and the operating systems that run most mobile devices – and we all benefit from superior products because of it.

I believe that AI will develop in a similar way. Today, several tech companies are developing leading closed models. But open source is quickly closing the gap. Last year, Llama 2 was only comparable to an older generation of models behind the frontier. This year, Llama 3 is competitive with the most advanced models and leading in some areas. Starting next year, we expect future Llama models to become the most advanced in the industry. But even before that, Llama is already leading on openness, modifiability, and cost efficiency.

Today we’re taking the next steps towards open source AI becoming the industry standard. We’re releasing Llama 3.1 405B, the first frontier-level open source AI model, as well as new and improved Llama 3.1 70B and 8B models. In addition to having significantly better cost/performance relative to closed models, the fact that the 405B model is open will make it the best choice for fine-tuning and distilling smaller models.

Beyond releasing these models, we’re working with a range of companies to grow the broader ecosystem. Amazon, Databricks, and NVIDIA are launching full suites of services to support developers fine-tuning and distilling their own models. Innovators like Groq have built low-latency, low-cost inference serving for all the new models. The models will be available on all major clouds including AWS, Azure, Google, Oracle, and more. Companies like Scale.AI, Dell, Deloitte, and others are ready to help enterprises adopt Llama and train custom models with their own data. As the community grows and more companies develop new services, we can collectively make Llama the industry standard and bring the benefits of AI to everyone.

Meta is committed to open source AI. I’ll outline why I believe open source is the best development stack for you, why open sourcing Llama is good for Meta, and why open source AI is good for the world and therefore a platform that will be around for the long term.

### Why Open Source AI Is Good for Developers

When I talk to developers, CEOs, and government officials across the world, I usually hear several themes:

1. **We need to train, fine-tune, and distill our own models.** Every organization has different needs that are best met with models of different sizes that are trained or fine-tuned with their specific data. On-device tasks and classification tasks require small models, while more complicated tasks require larger models. Now you’ll be able to take the most advanced Llama models, continue training them with your own data and then distill them down to a model of your optimal size – without us or anyone else seeing your data.

2. **We need to control our own destiny and not get locked into a closed vendor.** Many organizations don’t want to depend on models they cannot run and control themselves. They don’t want closed model providers to be able to change their model, alter their terms of use, or even stop serving them entirely. They also don’t want to get locked into a single cloud that has exclusive rights to a model. Open source enables a broad ecosystem of companies with compatible toolchains that you can move between easily.

3. **We need to protect our data.** Many organizations handle sensitive data that they need to secure and can’t send to closed models over cloud APIs. Other organizations simply don’t trust the closed model providers with their data. Open source addresses these issues by enabling you to run the models wherever you want. It is well-accepted that open source software tends to be more secure because it is developed more transparently.

4. **We need a model that is efficient and affordable to run.** Developers can run inference on Llama 3.1 405B on their own infra at roughly 50% the cost of using closed models like GPT-4o, for both user-facing and offline inference tasks.

5. **We want to invest in the ecosystem that’s going to be the standard for the long term.** Lots of people see that open source is advancing at a faster rate than closed models, and they want to build their systems on the architecture that will give them the greatest advantage long term.

### Why Open Source AI Is Good for Meta

Meta’s business model is about building the best experiences and services for people. To do this, we must ensure that we always have access to the best technology, and that we’re not locking into a competitor’s closed ecosystem where they can restrict what we build.

One of my formative experiences has been building our services constrained by what Apple will let us build on their platforms. Between the way they tax developers, the arbitrary rules they apply, and all the product innovations they block from shipping, it’s clear that Meta and many other companies would be freed up to build much better services for people if we could build the best versions of our products and competitors were not able to constrain what we could build. On a philosophical level, this is a major reason why I believe so strongly in building open ecosystems in AI and AR/VR for the next generation of computing.

People often ask if I’m worried about giving up a technical advantage by open sourcing Llama, but I think this misses the big picture for a few reasons:

1. **First**, to ensure that we have access to the best technology and aren’t locked into a closed ecosystem over the long term, Llama needs to develop into a full ecosystem of tools, efficiency improvements, silicon optimizations, and other integrations. If we were the only company using Llama, this ecosystem wouldn’t develop and we’d fare no better than the closed variants of Unix.

2. **Second**, I expect AI development will continue to be very competitive, which means that open sourcing any given model isn’t giving away a massive advantage over the next best models at that point in time. The path for Llama to become the industry standard is by being consistently competitive, efficient, and open generation after generation.

3. **Third**, a key difference between Meta and closed model providers is that selling access to AI models isn’t our business model. That means openly releasing Llama doesn’t undercut our revenue, sustainability, or ability to invest in research like it does for closed providers. (This is one reason several closed providers consistently lobby governments against open source.)

4. **Finally**, Meta has a long history of open source projects and successes. We’ve saved billions of dollars by releasing our server, network, and data center designs with Open Compute Project and having supply chains standardize on our designs. We benefited from the ecosystem’s innovations by open sourcing leading tools like PyTorch, React, and many more tools. This approach has consistently worked for us when we stick with it over the long term.

### Why Open Source AI Is Good for the World

I believe that open source is necessary for a positive AI future. AI has more potential than any other modern technology to increase human productivity, creativity, and quality of life – and to accelerate economic growth while unlocking progress in medical and scientific research. Open source will ensure that more people around the world have access to the benefits and opportunities of AI, that power isn’t concentrated in the hands of a small number of companies, and that the technology can be deployed more evenly and safely across society.

There is an ongoing debate about the safety of open source AI models, and my view is that open source AI will be safer than the alternatives. I think governments will conclude it’s in their interest to support open source because it will make the world more prosperous and safer.

My framework for understanding safety is that we need to protect against two categories of harm: unintentional and intentional. Unintentional harm is when an AI system may cause harm even when it was not the intent of those running it to do so. For example, modern AI models may inadvertently give bad health advice. Or, in more futuristic scenarios, some worry that models may unintentionally self-replicate or hyper-optimize goals to the detriment of humanity. Intentional harm is when a bad actor uses an AI model with the goal of causing harm.

It’s worth noting that unintentional harm covers the majority of concerns people have around AI – ranging from what influence AI systems will have on the billions of people who will use them to most of the truly catastrophic science fiction scenarios for humanity. On this front, open source should be significantly safer since the systems are more transparent and can be widely scrutinized. Historically, open source software has been more secure for this reason. Similarly, using Llama with its safety systems like Llama Guard will likely be safer and more secure than closed models. For this reason, most conversations around open source AI safety focus on intentional harm.

Our safety process includes rigorous testing and red-teaming to assess whether our models are capable of meaningful harm, with the goal of mitigating risks before release. Since the models are open, anyone is capable of testing for themselves as well. We must keep in mind that these models are trained by information that’s already on the internet, so the starting point when considering harm should be whether a model can facilitate more harm than information that can quickly be retrieved from Google or other search results.

When reasoning about intentional harm, it’s helpful to distinguish between what individual or small scale actors may be able to do as opposed to what large scale actors like nation states with vast resources may be able to do.

At some point in the future, individual bad actors may be able to use the intelligence of AI models to fabricate entirely new harms from the information available on the internet. At this point, the balance of power will be critical to AI safety. I think it will be better to live in a world where AI is widely deployed so that larger actors can check the power of smaller bad actors. This is how we’ve managed security on our social networks – our more robust AI systems identify and stop threats from less sophisticated actors who often use smaller scale AI systems. More broadly, larger institutions deploying AI at scale will promote security and stability across society. As long as everyone has access to similar generations of models – which open source promotes – then governments and institutions with more compute resources will be able to check bad actors with less compute.

The next question is how the US and democratic nations should handle the threat of states with massive resources like China. The United States’ advantage is decentralized and open innovation. Some people argue that we must close our models to prevent China from gaining access to them, but my view is that this will not work and will only disadvantage the US and its allies. Our adversaries are great at espionage, stealing models that fit on a thumb drive is relatively easy, and most tech companies are far from operating in a way that would make this more difficult. It seems most likely that a world of only closed models results in a small number of big companies plus our geopolitical adversaries having access to leading models, while startups, universities, and small businesses miss out on opportunities. Plus, constraining American innovation to closed development increases the chance that we don’t lead at all. Instead, I think our best strategy is to build a robust open ecosystem and have our leading companies work closely with our government and allies to ensure they can best take advantage of the latest advances and achieve a sustainable first-mover advantage over the long term.

When you consider the opportunities ahead, remember that most of today’s leading tech companies and scientific research are built on open source software. The next generation of companies and research will use open source AI if we collectively invest in it. That includes startups just getting off the ground as well as people in universities and countries that may not have the resources to develop their own state-of-the-art AI from scratch.

The bottom line is that open source AI represents the world’s best shot at harnessing this technology to create the greatest economic opportunity and security for everyone.

### Let’s Build This Together

With past Llama models, Meta developed them for ourselves and then released them, but didn’t focus much on building a broader ecosystem. We’re taking a different approach with this release. We’re building teams internally to enable as many developers and partners as possible to use Llama, and we’re actively building partnerships so that more companies in the ecosystem can offer unique functionality to their customers as well.

I believe the Llama 3.1 release will be an inflection point in the industry where most developers begin to primarily use open source, and I expect that approach to only grow from here. I hope you’ll join us on this journey to bring the benefits of AI to everyone in the world.

You can access the models now at [llama.meta.com](https://llama.meta.com).

💪,

http://mp.weixin.qq.com/s?__biz=MzA5MzIyMzI3Mw==&mid=2247488550&idx=3&sn=4d012791380dce49c8f1b69d7002f7de

知经Knowecon

北京大学经济学金融学统计学考研核心速通。

尼科尔森选择题题库一例：回顾赫克歇尔-俄林模型和斯托尔珀-萨缪尔森定理

做题的意义，训练的逻辑，为什么你的专业课必须考满分，以及 Don't be this fucking dog.

Elon Musk会是美国政治的新凯撒吗？回忆中间选民定理

尼科尔森选择题题库一例：回顾赫克歇尔-俄林模型和斯托尔珀-萨缪尔森定理

弗朗西斯·福山：自由主义在美国的衰败｜中英文对照

新晋诺奖得主达龙·阿西莫格鲁最新推文：特朗普治下的经济不容乐观｜中英文对照

Founder Mode / Paul Graham - 创始人模式 / 保罗·格雷厄姆

做题的意义，训练的逻辑，为什么你的专业课必须考满分，以及 Don't be this fucking dog.

DHH: Programmers Should Stop Celebrating Incompetence

反对学不会：用保姆级的统计学习题讲义和讲解

为什么我相信你是超人，统计学小小班新一周的交付计划，以及未来时间线超人版

交付了茆诗松5.1-5.3节的课程，以及人大802微观经济学部分2024-2007年的所有题目详解

周末思考：知识的开源协作、学习的工程化、人工智能

周末阅读：Paul Graham - 如何做伟大的工作，假设你是一个非常野心勃勃的人

我们今天和明天的交付安排

统计学小小班报名通道

通过仔细研究茆诗松和李子奈，我发现这并不难｜统计学小小班正式开讲

北大金融考研统计学小小班课程的课程安排、价格保护策略和优惠措施

成功路上Play的一环：我们对北大金融硕士改考微观经济学+统计学的外生冲击的认识和对策

在知识的富矿里进一步开掘

微观经济学接下来的核心工作是刷题：接下来的刷题安排0925-1110

中学数学能力对于经济学考研蛮重要的：人大802经济学2020年考研真题一例

微观经济学小小班同学开始每天刷一份考研真题，从人大802经济学开始

LSE EC411 微观经济学 Final 2019 Q5 博弈论问题详解

为什么我们的T恤要印上 Carthago Delenda Est：迦太基必须被毁灭

大促：购买Tee的同学可获赠价值51元的《微观经济学核心速通》一本

范里安《微观经济学：现代观点》学习笔记即将更新勘误

买Tee送书：获赠价值51元的《微观经济学核心速通》一本

小小班交付笔记：微观经济学更新006尼科尔森和103尼科尔森习题参考答案手写版

2024经验贴#12｜小语种生的北大汇丰考研一战高分备考攻略 - 政治、英语、数学、专业课全揭秘

2024经验贴#13｜工科同学一战跨考北大汇丰金融硕士经验分享

一份恰到好处的范里安《微观经济学：现代观点》学习笔记

小小班交付笔记：微观经济学更新006尼科尔森和范里安综合精讲和306拍卖理论专题

大促：购买Tee的同学可获赠价值51元的《微观经济学核心速通》一本

一份恰到好处的范里安《微观经济学：现代观点》学习笔记

KNOWECON ECONCLUB T-Shirt 2024Fall

小小班交付笔记：经济学辅导课程皆有更新，同时开源了三个脚本工具

2024经验贴#13｜工科同学一战跨考北大汇丰金融硕士经验分享

2024经验贴#12｜小语种生的北大汇丰考研一战高分备考攻略 - 政治、英语、数学、专业课全揭秘

2024经验贴#11｜一战上岸北大汇丰金融硕士经验分享：DOS AND DONTS

2024经验贴#10｜五千字长文分享北大软微金融科技备考经验

2024经验贴#9｜北大汇丰金融硕士备考专业课经验：原则、重点和计划

2024经验贴#7｜双非院校同学二战北大汇丰经济学硕士考研初试准备经验

2024经验贴#6｜巨高分｜985院校土木+金融背景二战北大光华考研初试准备经验

一份恰到好处的范里安《微观经济学：现代观点》学习笔记

Open Source AI Is the Path Forward

公众号中的英文文章也可以点击查词了？来测试一下我们开发的新能力

分类

时事

民生

政务

教育

文化

科技

财富

体娱

健康

情感

旅行

百科

职场

楼市

企业

乐活

学术

汽车

时尚

创业

美食

幽默

美体

文摘

原创标签

时事社会财经军事教育体育科技汽车科学房产搞笑综艺明星音乐动漫游戏时尚健康旅游美食生活摄影宠物职场育儿情感小说曲艺文化历史三农文学娱乐电影视频图片新闻宗教电视剧纪录片广告创意壁纸头像心灵鸡汤星座命理教育培训艺术文化金融财经健康医疗美妆时尚餐饮美食母婴育儿社会新闻工业农业时事政治星座占卜幽默笑话独立短篇连载作品文化历史科技互联网

发布位置

广东北京山东江苏河南浙江山西福建河北上海四川陕西湖南安徽湖北内蒙古江西云南广西甘肃辽宁黑龙江贵州新疆重庆吉林天津海南青海宁夏西藏香港澳门台湾美国加拿大澳大利亚日本新加坡英国西班牙新西兰韩国泰国法国德国意大利缅甸菲律宾马来西亚越南荷兰柬埔寨俄罗斯巴西智利卢森堡芬兰瑞典比利时瑞士土耳其斐济挪威朝鲜尼日利亚阿根廷匈牙利爱尔兰印度老挝葡萄牙乌克兰印度尼西亚哈萨克斯坦塔吉克斯坦希腊南非蒙古奥地利肯尼亚加纳丹麦津巴布韦埃及坦桑尼亚捷克阿联酋安哥拉