Tagged articles

Open source AI

115 articles · Page 2 of 2
Java Tech Enthusiast
Java Tech Enthusiast
Jul 12, 2024 · Artificial Intelligence

Why Alibaba’s Qwen‑2 Is Outperforming Global LLMs and What It Means for AI

After OpenAI halted API access in China, Alibaba’s Tongyi Qwen‑2 quickly rose to the top of global open‑source LLM leaderboards, surpassing Meta’s Llama‑3 and other contenders, with detailed benchmark scores, performance gains over previous versions, and implications for China’s AI ecosystem.

AI BenchmarkAlibabaChina AI
0 likes · 5 min read
Why Alibaba’s Qwen‑2 Is Outperforming Global LLMs and What It Means for AI
IT Services Circle
IT Services Circle
Jun 9, 2024 · Artificial Intelligence

Plagiarism Allegations Between Stanford's Llama3‑V and China's MiniCPM‑Llama3‑V 2.5 Model

The article details the controversy surrounding Stanford's Llama3‑V team admitting to copying the architecture and code of the Chinese MiniCPM‑Llama3‑V 2.5 model, presents new evidence of weight similarity, compares performance metrics, and discusses broader concerns about the recognition of Chinese AI research in the open‑source community.

AI ethicsLlama3-VMiniCPM
0 likes · 9 min read
Plagiarism Allegations Between Stanford's Llama3‑V and China's MiniCPM‑Llama3‑V 2.5 Model
21CTO
21CTO
May 28, 2024 · Artificial Intelligence

13 Open‑Source AI Projects That Made the 2024 GitHub Accelerator – A Deep Dive

This article showcases the 13 award‑winning open‑source AI projects featured in the 2024 GitHub Accelerator, highlighting each project's purpose, founders, key technologies, and how they advance machine‑learning, model training, deployment, and innovative AI applications across various domains.

AI ToolsGitHub AcceleratorLLM
0 likes · 9 min read
13 Open‑Source AI Projects That Made the 2024 GitHub Accelerator – A Deep Dive
NewBeeNLP
NewBeeNLP
Apr 22, 2024 · Artificial Intelligence

Why LLAMA‑3’s Scaling Laws Signal the Next AI Frontier

The article analyzes LLAMA‑3’s architectural tweaks, massive data expansion, scaling‑law implications, open‑source versus closed‑source dynamics, and the critical role of synthetic data in sustaining large‑model progress beyond 2025.

LLAMA-3Large Language ModelsOpen source AI
0 likes · 10 min read
Why LLAMA‑3’s Scaling Laws Signal the Next AI Frontier
21CTO
21CTO
Feb 29, 2024 · Artificial Intelligence

StarCoder2 Unveiled: Open-Source LLM That Outperforms Its Predecessor with Fewer Parameters

StarCoder2, the latest open-source large language model from ServiceNow, Hugging Face, and NVIDIA, offers three sizes—30B, 70B, and 150B parameters—delivering performance comparable to the original 150B StarCoder while being more efficient and freely accessible under the BigCode Open RAIL‑M license.

Artificial IntelligenceLLMOpen source AI
0 likes · 4 min read
StarCoder2 Unveiled: Open-Source LLM That Outperforms Its Predecessor with Fewer Parameters
DataFunSummit
DataFunSummit
Oct 27, 2023 · Artificial Intelligence

ChatGPT Technology, Domesticization Attempts, and Open‑Source Large Models

This article reviews the evolution and challenges of ChatGPT technology, describes the authors' efforts to localize and commercialize the model for the Chinese market, and introduces their open‑source Chinese large‑model initiative, including training methods, performance gaps, and future improvement directions.

ChatGPTChinese NLPLarge Language Models
0 likes · 11 min read
ChatGPT Technology, Domesticization Attempts, and Open‑Source Large Models
Baobao Algorithm Notes
Baobao Algorithm Notes
Jul 19, 2023 · Artificial Intelligence

Llama 2’s Breakthroughs: Architecture, Data, and Training Tricks Explained

Llama 2 advances open‑source large‑model research by expanding context length to 4096, adopting GQA attention, scaling training data to 2 trillion tokens, and introducing refined SFT and RLHF techniques such as Ghost Attention, margin‑based reward modeling, and iterative rejection sampling, all detailed in Meta’s 76‑page report.

Llama 2Open source AIRLHF
0 likes · 8 min read
Llama 2’s Breakthroughs: Architecture, Data, and Training Tricks Explained
DataFunSummit
DataFunSummit
May 17, 2023 · Artificial Intelligence

OpenAI Announces Plans to Release a New Open‑Source Large Language Model

OpenAI is set to launch its first open‑source large language model in four years, sparking debate over how this move could reshape the competitive landscape of AI, affect models like LLaMA, and intensify the open‑source versus closed‑source rivalry with Google.

AI competitionArtificial IntelligenceOpen source AI
0 likes · 6 min read
OpenAI Announces Plans to Release a New Open‑Source Large Language Model
Programmer DD
Programmer DD
Apr 18, 2023 · Artificial Intelligence

Can OpenAssistant Rival ChatGPT? Inside the Largest Open‑Source AI Assistant

This article examines OpenAssistant, the world’s largest open‑source ChatGPT replica, detailing its dataset of over 160 k annotated conversations, the fine‑tuned LLaMA and Pythia models, evaluation results against GPT‑3.5‑turbo, practical usage examples, and the project's current limitations and future directions.

AI datasetChatGPT alternativeOpen source AI
0 likes · 11 min read
Can OpenAssistant Rival ChatGPT? Inside the Largest Open‑Source AI Assistant
DataFunTalk
DataFunTalk
Feb 20, 2023 · Artificial Intelligence

ChatGPT Technology, Localization Efforts, and Open‑Source Large Models – Overview and Practices

This article presents an overview of ChatGPT technology, its evolution, current challenges, a three‑stage learning process, data organization and evaluation, details of domestic localization efforts, practical solutions, and the release of a Chinese open‑source large model with training guidance.

ChatGPTModel LocalizationOpen source AI
0 likes · 12 min read
ChatGPT Technology, Localization Efforts, and Open‑Source Large Models – Overview and Practices
21CTO
21CTO
Dec 30, 2022 · Artificial Intelligence

How a Chinese Developer Recreated ChatGPT with Google’s PaLM and RLHF

A Chinese engineer reverse‑engineered ChatGPT by building on Google’s massive PaLM model and applying reinforcement learning from human feedback, revealing the technical steps, challenges, and community reactions to this ambitious open‑source AI project.

ChatGPTOpen source AIPaLM
0 likes · 6 min read
How a Chinese Developer Recreated ChatGPT with Google’s PaLM and RLHF
Baidu Tech Salon
Baidu Tech Salon
Sep 2, 2022 · Artificial Intelligence

WAIC 2022: AI Open Source and Industrial Intelligence Summit Highlights China's AI Ecosystem Development

At the WAIC 2022 AI Open Source and Industrial Intelligence Summit in Shanghai, Baidu’s CTO outlined a TSMC‑like model for large‑scale AI, academicians highlighted intelligent vehicle connectivity and open‑source leadership, a new deep‑learning transformation base was unveiled, and PaddlePaddle’s 4.77 million developers underscored China’s rapidly expanding AI ecosystem across industry.

Artificial IntelligenceBaiduChina AI ecosystem
0 likes · 6 min read
WAIC 2022: AI Open Source and Industrial Intelligence Summit Highlights China's AI Ecosystem Development
21CTO
21CTO
Jul 9, 2022 · Artificial Intelligence

Meta Unveils NLLB-200: Open‑Source AI Model Translating 200 Languages

Meta has open‑sourced its new NLLB‑200 model, a single AI system that translates 200 languages with up to 44 % higher quality than its predecessor, supporting numerous low‑resource languages and powering billions of daily translations across Facebook and Instagram to improve user experience and content safety.

MetaNLLB-200Open source AI
0 likes · 3 min read
Meta Unveils NLLB-200: Open‑Source AI Model Translating 200 Languages