Artificial Intelligence 12 min read

Applying Large Language Models to Zhihu's Bridge Platform: Use Cases, Challenges, and Solutions

This article details how Zhihu's internal Bridge platform integrates large language models for business analysis, knowledge taxonomy, natural‑language‑to‑filter conversion, and ad‑hoc data queries, describing the workflow, technical hurdles, iterative improvements, and future directions.

DataFunTalk

Aug 19, 2023

Applying Large Language Models to Zhihu's Bridge Platform: Use Cases, Challenges, and Solutions

Introduction – The Bridge platform is Zhihu's internal operations analysis system. The article shares how large language models (LLMs) are applied to automate knowledge organization, natural‑language filtering, and data analysis, discussing both business impact and technical challenges.

1. Business Status and Background – Bridge supports six core scenarios: finding people, content, monitoring, opportunity discovery, and problem investigation. LLMs are used to combine external news‑driven inspiration with internal content‑driven inspiration, forming the first stage of knowledge taxonomy construction.

2. Knowledge System Classification – Two business forms are addressed: event aggregation from external news and content sedimentation into hierarchical taxonomies. LLMs assist in extracting key information, performing multi‑round clustering, naming clusters, and merging similar events, while a MapReduce‑like pipeline mitigates max‑token limits and workflow complexity. Advantages include automatic event naming and higher accuracy.

Event aggregation pipeline: news vectorization → high‑precision clustering → naming via LLM → hierarchical merging → final event generation.

Knowledge organization pipeline: content splitting → map phase (generate category names) → reduce phase (merge categories) → recursive merging until convergence.

Key solutions for token limits, clustering accuracy, and process complexity involve hierarchical clustering of LLM‑generated events, prompt size reduction, and a MapReduce‑style parallel framework.

3. Natural Language to Filter Conditions – Targets packaging, person, and content search. The task involves many filter criteria and complex logical combinations. The team fine‑tuned LLMs across four data‑construction stages (atomic conditions, logical combinations, fuzzy statements, and error‑prone cases) and iterated three model versions to address JSON truncation, output duplication, and logical errors.

Improvements included expanding token limits, random sampling to avoid repetitive outputs, and extensive JSON sample generation to correct format issues.

4. Natural Language Data Analysis – Focuses on ad‑hoc SQL generation from user queries. Challenges include handling diverse NL inputs, mapping queries to appropriate data sources, and ensuring business‑specific results. The solution uses dynamic prompts with FAISS‑based similarity search (MMR) to retrieve relevant examples, construct concise prompts respecting token limits, and generate SQL statements.

Online results show reduced cost, higher user adoption, and streamlined onboarding for new users, though accuracy remains a concern for complex queries, prompting future fine‑tuning efforts.

5. Summary and Outlook – The experience highlights pain points such as lack of mature PE practices, max‑token constraints, prompt engineering trial‑and‑error, model latency, and insufficient frameworks for large‑scale scenarios. Future directions include building dedicated frameworks for complex LLM tasks, leveraging business imagination to expand model capabilities, and continued fine‑tuning.

6. Q&A – Addresses model selection for event aggregation, evaluation methods, termination criteria for knowledge organization, and practical tips for improving LLM performance in production.

Original Source

Signed-in readers can open the original source through BestHub's protected redirect.

Republication Notice

This article has been distilled and summarized from source material, then republished for learning and reference. If you believe it infringes your rights, please contactand we will review it promptly.

Prompt engineering Large Language Models model fine-tuning natural language processing AI for business analytics knowledge taxonomy

Written by

DataFunTalk

Dedicated to sharing and discussing big data and AI technology applications, aiming to empower a million data scientists. Regularly hosts live tech talks and curates articles on big data, recommendation/search algorithms, advertising algorithms, NLP, intelligent risk control, autonomous driving, and machine learning/deep learning.

0 followers

Reader feedback

How this landed with the community

Rate this article

Was this worth your time?

Discussion

0 Comments

Thoughtful readers leave field notes, pushback, and hard-won operational detail here.