Average Ratings 1 Rating

Total
ease
features
design
support

Average Ratings 0 Ratings

Total
ease
features
design
support

No User Reviews. Be the first to provide a review:

Write a Review

Description

Qwen LLM represents a collection of advanced large language models created by Alibaba Cloud's Damo Academy. These models leverage an extensive dataset comprising text and code, enabling them to produce human-like text, facilitate language translation, craft various forms of creative content, and provide informative answers to queries. Key attributes of Qwen LLMs include: A range of sizes: The Qwen series features models with parameters varying from 1.8 billion to 72 billion, catering to diverse performance requirements and applications. Open source availability: Certain versions of Qwen are open-source, allowing users to access and modify the underlying code as needed. Multilingual capabilities: Qwen is equipped to comprehend and translate several languages, including English, Chinese, and French. Versatile functionalities: In addition to language generation and translation, Qwen models excel in tasks such as answering questions, summarizing texts, and generating code, making them highly adaptable tools for various applications. Overall, the Qwen LLM family stands out for its extensive capabilities and flexibility in meeting user needs.

Description

Sky-T1-32B-Preview is an innovative open-source reasoning model crafted by the NovaSky team at UC Berkeley's Sky Computing Lab. It delivers performance comparable to proprietary models such as o1-preview on various reasoning and coding assessments, while being developed at a cost of less than $450, highlighting the potential for budget-friendly, advanced reasoning abilities. Fine-tuned from Qwen2.5-32B-Instruct, the model utilized a meticulously curated dataset comprising 17,000 examples spanning multiple fields, such as mathematics and programming. The entire training process was completed in just 19 hours using eight H100 GPUs with DeepSpeed Zero-3 offloading technology. Every component of this initiative—including the data, code, and model weights—is entirely open-source, allowing both academic and open-source communities to not only replicate but also improve upon the model's capabilities. This accessibility fosters collaboration and innovation in the realm of artificial intelligence research and development.

API Access

Has API

API Access

Has API

Screenshots View All

No images available

Screenshots View All

Integrations

AiAssistWorks
Alibaba Cloud
Athene-V2
Axolotl
Decompute Blackbird
Featherless
Hugging Face
LLaMA-Factory
LM-Kit.NET
ModelScope
Oumi
Qwen Chat
SambaNova
Symflower
TypeThink
WebLLM
Zemith

Integrations

AiAssistWorks
Alibaba Cloud
Athene-V2
Axolotl
Decompute Blackbird
Featherless
Hugging Face
LLaMA-Factory
LM-Kit.NET
ModelScope
Oumi
Qwen Chat
SambaNova
Symflower
TypeThink
WebLLM
Zemith

Pricing Details

Free
Free Trial
Free Version

Pricing Details

Free
Free Trial
Free Version

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Deployment

Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Customer Support

Business Hours
Live Rep (24/7)
Online Support

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Types of Training

Training Docs
Webinars
Live Training (Online)
In Person

Vendor Details

Company Name

Alibaba

Founded

1999

Country

China

Website

github.com/QwenLM/Qwen

Vendor Details

Company Name

NovaSky

Country

United States

Website

novasky-ai.github.io/posts/sky-t1/

Product Features

Product Features

Alternatives

Phi-3 Reviews

Phi-3

Microsoft

Alternatives

OLMo 2 Reviews

OLMo 2

Ai2
Qwen2-VL Reviews

Qwen2-VL

Alibaba
Open R1 Reviews

Open R1

Open R1
Qwen2 Reviews

Qwen2

Alibaba
Smaug-72B Reviews

Smaug-72B

Abacus