Tagged articles

large language model

831 articles · Page 9 of 9
Architecture Digest
Architecture Digest
Mar 17, 2023 · Artificial Intelligence

Baidu’s Ernie Bot (Wenxin Yiyan) vs GPT‑4: Capabilities, Technical Foundations, and Market Reaction

The article reviews Baidu's launch of the multimodal large language model Wenxin Yiyan, compares its literary, business, mathematical, Chinese‑understanding and multimodal abilities with GPT‑4, explains the underlying six‑core technologies and hardware stack, and reports the mixed market and netizen response.

AIBaiduErnie Bot
0 likes · 11 min read
Baidu’s Ernie Bot (Wenxin Yiyan) vs GPT‑4: Capabilities, Technical Foundations, and Market Reaction
Smart Era Software Development
Smart Era Software Development
Mar 17, 2023 · Artificial Intelligence

Wenxin Yiyan vs GPT-4: Live Demo Shows Baidu’s New AI Model in Action

The article presents a side‑by‑side demonstration of Baidu’s newly released Wenxin Yiyan and OpenAI’s GPT‑4 across literary creation, business copywriting, mathematical reasoning, Chinese idiom interpretation, acrostic poetry, and multimodal generation, then explains the underlying six‑core technologies and Baidu’s hardware‑cloud strategy while reporting audience reactions.

Chinese NLPERNIEGPT-4
0 likes · 11 min read
Wenxin Yiyan vs GPT-4: Live Demo Shows Baidu’s New AI Model in Action
21CTO
21CTO
Mar 15, 2023 · Artificial Intelligence

What Makes OpenAI’s New GPT‑4 a Game‑Changer for Multimodal AI?

OpenAI’s GPT‑4, a multimodal large language model that accepts text and image inputs, powers ChatGPT and Bing, offers improved creativity and problem‑solving while still facing hallucination risks, and is now available via ChatGPT Plus and an open API for developers.

AI safetyGPT-4Multimodal AI
0 likes · 5 min read
What Makes OpenAI’s New GPT‑4 a Game‑Changer for Multimodal AI?
DataFunSummit
DataFunSummit
Mar 15, 2023 · Artificial Intelligence

Key Features and Capabilities of OpenAI's GPT‑4

OpenAI's GPT‑4, a large multimodal language model, expands token limits, adds image understanding, demonstrates strong reasoning on professional exams, supports many languages, and is already integrated into Microsoft Bing, while offering various access options and improved safety compared to its predecessor.

AIGPT-4Microsoft Bing
0 likes · 9 min read
Key Features and Capabilities of OpenAI's GPT‑4
Tencent Advertising Technology
Tencent Advertising Technology
Mar 2, 2023 · Artificial Intelligence

Tencent's HunYuan‑NLP 1T Large‑Scale AI Model: Training Techniques, Optimization, and Real‑World Applications

This article details Tencent's development of the 1‑trillion‑parameter HunYuan‑NLP model, covering its MoE architecture, cost‑effective pre‑training strategies, distributed training framework, model compression toolkit, and successful deployment across advertising, gaming, and other Tencent services.

AI infrastructureMixture of Expertslarge language model
0 likes · 17 min read
Tencent's HunYuan‑NLP 1T Large‑Scale AI Model: Training Techniques, Optimization, and Real‑World Applications
DataFunSummit
DataFunSummit
Feb 26, 2023 · Artificial Intelligence

Fudan University's MOSS: China's First Conversational Large Language Model

Fudan University's Natural Language Processing Lab introduced MOSS, the country's first conversational large language model capable of dialogue generation, programming, factual QA and ethical reasoning, with plans for open‑source release despite current limitations in Chinese language proficiency.

AIFudan UniversityMOSS
0 likes · 3 min read
Fudan University's MOSS: China's First Conversational Large Language Model
DataFunSummit
DataFunSummit
Feb 24, 2023 · Artificial Intelligence

Baidu PLATO Open‑Domain Dialogue Model: Technology, Challenges, and Applications

The article presents Baidu's PLATO open‑domain dialogue system, detailing its evolution from expert‑rule to retrieval‑based and large‑scale generative models, describing its hidden‑variable architecture, major research challenges such as persona stability, long‑term memory, knowledge accuracy, and showcasing real‑world applications and Q&A from a DataFunSummit2022 livestream.

AIKnowledge RetrievalLong-Term Memory
0 likes · 25 min read
Baidu PLATO Open‑Domain Dialogue Model: Technology, Challenges, and Applications
Programmer DD
Programmer DD
Feb 21, 2023 · Artificial Intelligence

Meet MOSS: China’s Homegrown ChatGPT Rival and Its Capabilities

MOSS, a Chinese large‑language model released by Fudan University, offers ChatGPT‑like functions such as text generation, summarization, translation, and code writing, while being open‑source and free during preview, yet it still lags behind due to limited data, compute, and model size.

AIChatGPTFudan University
0 likes · 11 min read
Meet MOSS: China’s Homegrown ChatGPT Rival and Its Capabilities
DataFunTalk
DataFunTalk
Feb 20, 2023 · Artificial Intelligence

ChatGPT Technology, Localization Efforts, and Open‑Source Large Models – Overview and Practices

This article presents an overview of ChatGPT technology, its evolution, current challenges, a three‑stage learning process, data organization and evaluation, details of domestic localization efforts, practical solutions, and the release of a Chinese open‑source large model with training guidance.

ChatGPTModel Localizationdata annotation
0 likes · 12 min read
ChatGPT Technology, Localization Efforts, and Open‑Source Large Models – Overview and Practices
DataFunTalk
DataFunTalk
Feb 19, 2023 · Artificial Intelligence

How ChatGPT Works: An In‑Depth Explanation by Stephen Wolfram

This article provides a comprehensive, step‑by‑step explanation of how ChatGPT generates text, covering token probabilities, n‑gram models, embeddings, attention mechanisms, and the Transformer architecture, while illustrating concepts with Wolfram‑language examples and visualizations.

AIChatGPTNeural Network
0 likes · 20 min read
How ChatGPT Works: An In‑Depth Explanation by Stephen Wolfram
Tencent Cloud Developer
Tencent Cloud Developer
Feb 14, 2023 · Artificial Intelligence

ChatGPT: Technology, Impact, and Future Perspectives

Since its November 2022 launch, OpenAI’s ChatGPT—built on Transformer‑based generative AI—has surged to over 100 million users, demonstrated capabilities from MBA exams to software‑engineer interviews, sparked a multibillion‑dollar market with paid subscriptions and Microsoft investment, spurred rival models like Claude, and is reshaping human‑computer interaction while raising ethical concerns and promising multimodal, industry‑specific future applications.

ChatGPTTransformergenerative AI
0 likes · 15 min read
ChatGPT: Technology, Impact, and Future Perspectives
Open Source Linux
Open Source Linux
Feb 10, 2023 · Artificial Intelligence

What Makes ChatGPT Tick? Features, Architecture, Limits, and Future Opportunities

This article provides a comprehensive overview of ChatGPT, covering its origins within OpenAI, core features, underlying GPT‑3.5 architecture, reinforcement learning from human feedback, current limitations, and future directions such as model compression, RLAIF, and expanding industry applications.

AIGCArtificial IntelligenceChatGPT
0 likes · 20 min read
What Makes ChatGPT Tick? Features, Architecture, Limits, and Future Opportunities
21CTO
21CTO
Feb 7, 2023 · Artificial Intelligence

Google’s Bard vs ChatGPT: Inside the New AI Chatbot and Its LaMDA Roots

Google unveiled its new conversational AI, Bard, built on the LaMDA model, positioning it as a direct competitor to ChatGPT; the article details its public testing, technical foundations, feature set, and key differences such as real‑time web integration and cost‑free access.

AI chatbotChatGPTGoogle AI
0 likes · 11 min read
Google’s Bard vs ChatGPT: Inside the New AI Chatbot and Its LaMDA Roots
Architect
Architect
Dec 20, 2022 · Artificial Intelligence

Understanding ChatGPT: Architecture, Training Process, Features, and Applications

An in‑depth overview of ChatGPT covering its conversational model nature, core technologies such as InstructGPT, large language model capabilities, RLHF training pipeline, strengths, limitations, safety mechanisms, and potential applications across content creation, search, and multimodal integration.

ApplicationsArtificial IntelligenceChatGPT
0 likes · 19 min read
Understanding ChatGPT: Architecture, Training Process, Features, and Applications
Architecture Digest
Architecture Digest
Dec 15, 2022 · Artificial Intelligence

Technical Overview of ChatGPT: Training Pipeline, RLHF, and Its Potential to Replace Search Engines

This article explains ChatGPT's underlying technology—including its three‑stage training pipeline with supervised fine‑tuning, reward‑model learning, and reinforcement learning from human feedback—while analyzing whether the model can realistically replace traditional search engines such as Google or Baidu.

AIChatGPTRLHF
0 likes · 15 min read
Technical Overview of ChatGPT: Training Pipeline, RLHF, and Its Potential to Replace Search Engines
IT Architects Alliance
IT Architects Alliance
Dec 13, 2022 · Artificial Intelligence

Technical Principles and Training Process of ChatGPT

The article explains ChatGPT’s underlying technology, detailing its three-stage training pipeline—supervised fine‑tuning, reward‑model learning, and reinforcement learning with PPO—while discussing its strengths, limitations, and potential integration with traditional search engines.

AIChatGPTLLM
0 likes · 14 min read
Technical Principles and Training Process of ChatGPT
Tencent Cloud Developer
Tencent Cloud Developer
Dec 9, 2022 · Artificial Intelligence

An Overview of ChatGPT: Technology, Training Process, and Applications

The article outlines ChatGPT’s conversational capabilities, its InstructGPT‑based architecture, a three‑stage RLHF training pipeline involving supervised fine‑tuning, human‑ranked response generation, and PPO optimization, and discusses its strengths, limitations, diverse applications, and future directions for multimodal, up‑to‑date assistants.

AI ApplicationsChatGPTPPO
0 likes · 18 min read
An Overview of ChatGPT: Technology, Training Process, and Applications
Architect's Guide
Architect's Guide
Dec 9, 2022 · Artificial Intelligence

Technical Principles and Training Process of ChatGPT

The article explains how ChatGPT builds on the GPT‑3.5 large language model, using human‑annotated data and Reinforcement Learning from Human Feedback (RLHF) across three training stages to improve instruction understanding, answer quality, and continual model enhancement, while also discussing its potential to complement or replace traditional search engines.

AIChatGPTInstruction Tuning
0 likes · 15 min read
Technical Principles and Training Process of ChatGPT
IT Architects Alliance
IT Architects Alliance
Dec 8, 2022 · Artificial Intelligence

Technical Principles and Training Process of ChatGPT

This article explains the technical foundations of ChatGPT, detailing its three-stage training pipeline—supervised fine‑tuning with human‑annotated data, reward model training via pairwise ranking, and reinforcement learning from human feedback—while also discussing its limitations compared to traditional search engines and potential future enhancements.

AIChatGPTRLHF
0 likes · 14 min read
Technical Principles and Training Process of ChatGPT
Top Architect
Top Architect
Dec 7, 2022 · Artificial Intelligence

Technical Principles of ChatGPT and Its Prospects for Replacing Traditional Search Engines

The article explains how ChatGPT builds on GPT‑3.5 with supervised fine‑tuning, reward‑model training and reinforcement learning from human feedback, analyzes why it cannot yet replace search engines due to hallucinations, knowledge freshness and cost, and proposes a hybrid architecture that combines LLM generation with traditional retrieval to overcome these limitations.

AIChatGPTRLHF
0 likes · 16 min read
Technical Principles of ChatGPT and Its Prospects for Replacing Traditional Search Engines
DataFunTalk
DataFunTalk
Oct 7, 2022 · Artificial Intelligence

Overview of Baidu's PLATO Open‑Domain Dialogue Technology, Challenges, and Applications

This article introduces Baidu's PLATO open‑domain dialogue technology, explains the evolution from rule‑based to retrieval‑based and large‑scale generative models, discusses major challenges such as persona stability, long‑term memory, knowledge accuracy, and proactive conversation, and showcases real‑world applications and Q&A insights.

AI ChallengesChatbot ApplicationsOpen-domain Dialogue
0 likes · 23 min read
Overview of Baidu's PLATO Open‑Domain Dialogue Technology, Challenges, and Applications
DataFunTalk
DataFunTalk
Jun 30, 2022 · Artificial Intelligence

OBERT: A Billion‑Parameter Pretrained Language Model for Large‑Scale NLP Applications

The OPPO XiaoBu team introduced OBERT, a series of 100M‑, 300M‑, and 1B‑parameter pretrained language models that leverage massive TB‑scale corpora, multi‑granular masking, retrieval‑augmented training, and distributed acceleration to achieve state‑of‑the‑art results on CLUE and KgCLUE benchmarks while enabling efficient industrial deployment.

Knowledge augmentationNLPfine-tuning
0 likes · 12 min read
OBERT: A Billion‑Parameter Pretrained Language Model for Large‑Scale NLP Applications
21CTO
21CTO
Jun 14, 2022 · Artificial Intelligence

Does Google’s LaMDA Really Possess Sentience? A Deep Dive into the Debate

The article examines the controversy surrounding Google’s LaMDA chatbot, detailing engineer Blake Lemoine’s claims of sentience, his suspension, the model’s technical specs, contrasting expert opinions from figures like Andrej Karpathy and Gary Marcus, and ultimately argues that LaMDA’s apparent emotions are a projection rather than true consciousness.

AI ethicsGoogleLaMDA
0 likes · 8 min read
Does Google’s LaMDA Really Possess Sentience? A Deep Dive into the Debate
Tencent Tech
Tencent Tech
Apr 29, 2022 · Artificial Intelligence

Tencent’s Hunyuan AI Model Tops CLUE Leaderboard with Record Score

Tencent’s Hunyuan AI large model shattered records by scoring 80.888 to claim first place on the CLUE benchmark, showcasing its advanced natural language processing, multimodal abilities, curriculum‑learning training approach, and real‑world deployments in WeChat Search and advertising.

AICLUEHunyuan
0 likes · 3 min read
Tencent’s Hunyuan AI Model Tops CLUE Leaderboard with Record Score
DataFunTalk
DataFunTalk
Sep 22, 2021 · Artificial Intelligence

Baidu Unveils PLATO-XL: A 110‑Billion‑Parameter Bilingual Dialogue Generation Model

Baidu's newly released PLATO‑XL, a 110‑billion‑parameter bilingual pre‑training dialogue model, surpasses previous large‑scale models, introduces multi‑role awareness for consistent multi‑turn conversations, and demonstrates state‑of‑the‑art performance across open‑domain, knowledge‑grounded, and task‑oriented dialogue tasks.

Natural Language ProcessingPLATO-XLbilingual AI
0 likes · 9 min read
Baidu Unveils PLATO-XL: A 110‑Billion‑Parameter Bilingual Dialogue Generation Model
DataFunTalk
DataFunTalk
Jul 8, 2021 · Artificial Intelligence

Baidu ERNIE 3.0: Knowledge‑Enhanced 100B‑Parameter Model Sets New Chinese NLP Benchmarks and Tops SuperGLUE

Baidu's ERNIE 3.0 introduces a 100‑billion‑parameter, knowledge‑graph‑augmented language model that breaks 54 Chinese NLP benchmarks, achieves human‑level performance on SuperGLUE, and demonstrates strong generation and zero‑shot capabilities, now available for public demo and research.

BaiduERNIE 3.0Knowledge Graph
0 likes · 7 min read
Baidu ERNIE 3.0: Knowledge‑Enhanced 100B‑Parameter Model Sets New Chinese NLP Benchmarks and Tops SuperGLUE