Alibaba Cloud Open-Sources More LLMs with Diverse Sizes and Multimodal Features – More Major Contributions to Open-Source Community

Hangzhou, China – Alibaba Cloud, the digital technology and intelligence backbone of Alibaba Group, announced that it has open-sourced two large language models (LLM), Qwen-72B and Qwen-1.8B, the 72-billion-parameter and 1.8-billion-parameter versions of its proprietary foundation model Tongyi Qianwen, on its AI model community ModelScope, and the collaborative AI platform Hugging Face.

In addition, Alibaba Cloud makes available more multimodal LLMs including Qwen-Audio and Qwen-Audio-Chat, a pre-trained audio understanding model and its conversationally fine tuned version for research and commercial purposes.

As of today, the cloud computing pioneer has contributed various sizes of LLMs with parameters ranging from 1.8B, 7B, 14B to 72B, as well as multimodal LLMs with audio and visual understanding features.

“Building up an open-source ecosystem is critical to promoting the development of LLM and AI applications building. We aspire to become the most open cloud and make generative AI capabilities accessible to everyone,” said Jingren Zhou, CTO of Alibaba Cloud. “To achieve that goal, we’ll continue to share our cutting-edge technology and facilitate the development of the open-source community together with our partners.”

Pre-trained on over 3 trillion tokens, the 72-billion-parameter model outperforms other major open-source models in ten benchmarks, including Massive Multi-task Language Understanding (MMLU) benchmark that measures the model’s multitask accuracy, HumanEval that tests code generation capabilities and GSM8K, a benchmark for arithmetic problems, to name a few.

Qwen-72B outperforms other major open-source models in ten benchmarks

The model also exhibits proficiency in tackling a variety of intricate tasks, including role-playing and language style transfer, referring to the ability of the LLM to assume a specific role or persona and generate more contextually relevant responses consistent with the persona. Such features can be useful in AI applications such as personalized chatbots.

Companies and research institutions can access the Qwen-72B model’s code, model weights, and documentation and use them for free for research purposes. For commercial uses, the models will be free to use for companies with fewer than 100 million monthly active users.

Alibaba Cloud also announced that it has open-sourced the 1.8-billion-parameter of its LLM that can run on the edge. The lightweight LLM enables inference on edge devices with constrained computational resource, making it possible to be deployed on end devices such as cellphones.

The smaller-sized version, with less computing resource requirement, can be useful for individuals looking for a more cost-effective, easy-to-deploy option in using LLMs. The 1.8B model is currently only available for research purposes.

To offer LLMs that can process a greater variety of input formats, Alibaba Cloud also announced that it has open-sourced Qwen-Audio and Qwen-Audio-Chat, the models with enhanced audio understanding capabilities for research and commercial purposes.

Qwen-Audio can understand text and audio input in diverse formats, including human speech, natural sound and music and produce text as output. It is capable of performing over 30 audio processing tasks, such as multi-language transcription, speech editing, audio caption analysis etc. Its conversationally fine tuned version, Qwen-Audio-Chat, can support multiple rounds of question-and-answering based on the audio and perform diverse audio-oriented tasks, such as detection of emotions and tones in human speeches.

The initiative marks another attempt from Alibaba Cloud to offer multi-modal large language models that can understand datatypes beyond text to the open-source community. Earlier this year, it announced the launch of open-source Large Vision Language Model Qwen-VL and its chat version Qwen-VL-Chat that can understand visual information and perform visual tasks.

The open-sourced LLM models, including Qwen-7B, Qwen-14B and Qwen-VL and their conversationally fine tuned versions, have gained a combined downloads of over 1.5 million times on Alibaba Cloud’s open-source AI model community ModelScope and Hugging Face since August. ModelScope has become the largest AI model community in China, boasting over 2.8 million active developers, with over 100 million model downloads to date.

For more information, please check out the details of Qwen-72B and Qwen-1.8B on ModelScope, Hugging Face and GitHub pages.

For a demo of Qwen-Audio, please click here: https://qwen-audio.github.io/Qwen-Audio/.

Liked this post? Follow SwirlingOverCoffee on Facebook, YouTube, and Instagram.

Follow Us

Trending News

Samsung Ridicules Apple’s Adoption of Big Screens on the iPhone 6 and 6 Plus

Walk like an Egyptian with Street View in Google Maps

ASUS launches Transformer Book Flip – From Laptop to Tablet in Seconds

Xiaomi Redmi 1s Goes on Sale on Sep 4 for P5,599

ASUS Memo Pad 7 Unboxing and First Impressions

Huawei Ascend G6 Review – Midrange Spec’d Selfie Phone

TDK 2-in-1 Micro USB 2.0 Flash Drive Review

ASUS ZenFone 4 Unboxing and First Impressions

ASUS ZenFone 5 Unboxing and First Impressions

Enjoy up to 40% discount on premium accessories when you buy a MacBook Air M3 at Power Mac Center

Coastal Grounds, brewing in Siargao

Radenta Brings Fun in Learning with Fable Robotics

Cool your summer with Home Credit’s abot-kayang inverter appliances

Showcase your moments of Summer Delight with OPPO

BSS (SEVENTEEN) Showcases the AI Eraser capabilities using the OPPO Reno11 F 5G

Blog Post

Advertisement

Advertisement

Instagram

Search

Follow Me On

Follow Us

Trending News

Samsung Ridicules Apple’s Adoption of Big Screens on the iPhone 6 and 6 Plus

Walk like an Egyptian with Street View in Google Maps

ASUS launches Transformer Book Flip – From Laptop to Tablet in Seconds

Xiaomi Redmi 1s Goes on Sale on Sep 4 for P5,599

ASUS Memo Pad 7 Unboxing and First Impressions

Huawei Ascend G6 Review – Midrange Spec’d Selfie Phone

TDK 2-in-1 Micro USB 2.0 Flash Drive Review

ASUS ZenFone 4 Unboxing and First Impressions

ASUS ZenFone 5 Unboxing and First Impressions

Enjoy up to 40% discount on premium accessories when you buy a MacBook Air M3 at Power Mac Center

Coastal Grounds, brewing in Siargao

Radenta Brings Fun in Learning with Fable Robotics

Cool your summer with Home Credit’s abot-kayang inverter appliances

Showcase your moments of Summer Delight with OPPO

BSS (SEVENTEEN) Showcases the AI Eraser capabilities using the OPPO Reno11 F 5G

Blog Post

Alibaba Cloud Open-Sources More LLMs with Diverse Sizes and Multimodal Features – More Major Contributions to Open-Source Community

SOC Staff

Related posts

Sophos Demonstrates How to Make ChatGPT a Cybersecurity Co-Pilot

Study of unsurveyed people shows Leni ahead of Bongbong

FIBA Basketball World Cup 2023 Opening Game goers get FREE Jollibee Cheesy Yumburgers

This Spooky Season, Tune In to the Tricks and Treats Lurking on Spotify

Home Credit offers smartphone treats at MemoXpress’ 24th Anniversary blowout

STT GDC Philippines cultivates future-ready talent to support data center expansion

Advertisement

Advertisement

Instagram