Tagged articles

AI research

312 articles · Page 3 of 4
AIWalker
AIWalker
Jan 17, 2025 · Artificial Intelligence

InternLM 3.0: Boosting Model Performance with Only 4 TB of Training Data

Shanghai AI Laboratory’s InternLM 3.0 upgrade demonstrates that refining data quality—measured as intelligence‑per‑token—can replace massive datasets, achieving higher reasoning and dialogue capabilities with just 4 TB of tokens, cutting training cost by over 75 % while approaching GPT‑4‑level performance.

AI researchData EfficiencyInternLM
0 likes · 9 min read
InternLM 3.0: Boosting Model Performance with Only 4 TB of Training Data
DevOps
DevOps
Jan 7, 2025 · Artificial Intelligence

Microsoft’s 2025 AI Predictions: Stronger Models, AI Agents, AI Companions, Efficient Resources, Testing & Customization, and Accelerated Scientific Research

Microsoft outlines six 2025 AI forecasts—including more powerful models, autonomous AI agents reshaping work, AI companions aiding daily life, greener resource use, rigorous testing and customization, and AI-driven scientific breakthroughs—highlighting how these advances will transform industries, research, and everyday experiences.

2025 predictionsAIAI models
0 likes · 8 min read
Microsoft’s 2025 AI Predictions: Stronger Models, AI Agents, AI Companions, Efficient Resources, Testing & Customization, and Accelerated Scientific Research
21CTO
21CTO
Jan 2, 2025 · Artificial Intelligence

2025 AI Breakthroughs: Unlimited Memory & Intelligent Agents, Says Eric Schmidt

Former Google CEO Eric Schmidt warns that AI is on the brink of a transformative era, highlighting three 2025 breakthroughs—unlimited context memory, autonomous AI agents, and text‑to‑action programming—while also stressing the looming risks of energy consumption, security threats, and the need for ethical safeguards.

AI memoryAI researchAI safety
0 likes · 14 min read
2025 AI Breakthroughs: Unlimited Memory & Intelligent Agents, Says Eric Schmidt
DaTaobao Tech
DaTaobao Tech
Dec 30, 2024 · Artificial Intelligence

AI Research Highlights: AAAI 2025 & NeurIPS 2024 Breakthroughs in Image Generation

This article compiles recent AI research breakthroughs presented at AAAI 2025 and NeurIPS 2024, summarizing eight papers on multi‑condition image generation, mixed auto‑regressive models, hallucination mitigation in vision‑language models, quantized diffusion denoising, facial part swapping, language‑guided concept vectors, attribution consistency, and video virtual try‑on, with links to each work.

AAAI 2025AI researchGenerative Models
0 likes · 13 min read
AI Research Highlights: AAAI 2025 & NeurIPS 2024 Breakthroughs in Image Generation
Baobao Algorithm Notes
Baobao Algorithm Notes
Dec 16, 2024 · Artificial Intelligence

What Do Leading Open‑Source LLMs Do After Pretraining? A Deep Dive into Post‑Training Strategies

This article surveys the post‑training pipelines of major open‑source large language models released this year, detailing their alignment algorithms, data synthesis, reward modeling, DPO/GRPO variants, long‑context handling, tool use, and model‑averaging techniques, and highlights emerging trends such as data‑centric pipelines and iterative weak‑to‑strong alignment.

AI researchAlignmentLLM
0 likes · 99 min read
What Do Leading Open‑Source LLMs Do After Pretraining? A Deep Dive into Post‑Training Strategies
Alipay Experience Technology
Alipay Experience Technology
Nov 27, 2024 · Artificial Intelligence

EchoMimicV2: High‑Quality Audio‑Driven Half‑Body Human Animation with Simple Inputs

EchoMimicV2 is an open‑source digital‑human framework that generates high‑quality half‑body animation videos from a single reference image, an audio clip, and a hand‑gesture sequence, addressing challenges of facial portrait limits, complex condition injection, and inference latency in audio‑driven animation.

AI researchVideo Generationaudio-driven animation
0 likes · 18 min read
EchoMimicV2: High‑Quality Audio‑Driven Half‑Body Human Animation with Simple Inputs
360 Tech Engineering
360 Tech Engineering
Nov 15, 2024 · Artificial Intelligence

Advances in Multimodal Large Models and Document Understanding Presented at the 2024 Global Machine Learning Conference (Beijing)

At the 2024 Global Machine Learning Conference in Beijing, 360 AI Research Institute showcased cutting‑edge multimodal large‑model research, fine‑grained open‑world object detection, and document understanding technologies, highlighting open‑source releases, real‑world deployments, and competitive achievements in AI competitions.

AI researchKnowledge Graphdocument understanding
0 likes · 7 min read
Advances in Multimodal Large Models and Document Understanding Presented at the 2024 Global Machine Learning Conference (Beijing)
Tencent Cloud Developer
Tencent Cloud Developer
Nov 6, 2024 · Artificial Intelligence

Overview of Tencent Hunyuan Large and 3D Generation Model Open‑Source Release

Tencent has open‑sourced its 389‑billion‑parameter Hunyuan Large Mixture‑of‑Experts model—featuring 52 B active parameters, 256 K token context, novel routing, KV‑cache compression, and advanced training optimizations that beat leading open‑source models—and its first text‑to‑3D/image‑to‑3D Hunyuan 3D Generation model, both downloadable via GitHub, Hugging Face, and Tencent Cloud.

3D GenerationAI researchMixture of Experts
0 likes · 9 min read
Overview of Tencent Hunyuan Large and 3D Generation Model Open‑Source Release
Meituan Technology Team
Meituan Technology Team
Oct 31, 2024 · Artificial Intelligence

Selected Meituan Papers from CIKM 2024: Summaries of Eight Research Works

This article highlights eight Meituan research papers accepted at CIKM 2024—spanning self‑supervised sequential recommendation, rating‑consistent explanation generation, CTR prediction via recommendation pre‑training, cross‑domain interest transfer, multimodal vector retrieval, design‑aware poster layout, order‑fulfillment cycle‑time forecasting, and delivery‑scope substitution—offering insights from both internal and university collaborations.

AI researchCTR predictionCross‑Domain Recommendation
0 likes · 16 min read
Selected Meituan Papers from CIKM 2024: Summaries of Eight Research Works
Baobao Algorithm Notes
Baobao Algorithm Notes
Oct 30, 2024 · Artificial Intelligence

How to Choose High-Quality Instruction Data for LLM Fine‑Tuning: Methods Compared

This article surveys and categorizes instruction data selection techniques for large language model fine‑tuning, explaining metric‑based, trainable‑LLM, powerful‑LLM, and small‑model approaches, detailing representative papers, their pipelines, and empirical findings on data quality and diversity.

AI researchInstruction TuningLLM data selection
0 likes · 15 min read
How to Choose High-Quality Instruction Data for LLM Fine‑Tuning: Methods Compared
AntTech
AntTech
Oct 29, 2024 · Artificial Intelligence

Three Ant Group Papers Featured at EMNLP 2024: Dynamic Transformers, Plug‑and‑Play Visual Reasoner, and Efficient Fine‑Tuning of Large Language Models

This announcement introduces three Ant Group papers accepted at EMNLP 2024—Mixture‑of‑Modules for dynamic Transformer assembly, a plug‑and‑play visual reasoning framework built via data synthesis, and a layer‑wise importance‑aware efficient fine‑tuning method for large language models—highlighting their innovations and upcoming live presentations.

AI researchEMNLP 2024Large Language Models
0 likes · 6 min read
Three Ant Group Papers Featured at EMNLP 2024: Dynamic Transformers, Plug‑and‑Play Visual Reasoner, and Efficient Fine‑Tuning of Large Language Models
Baobao Algorithm Notes
Baobao Algorithm Notes
Oct 24, 2024 · Artificial Intelligence

How NoteLLM-2 Boosts Multimodal Recommendations with In-Content Learning

NoteLLM-2 introduces multimodal In-Content Learning and Late Fusion to overcome visual‑modality bias in end‑to‑end fine‑tuned large representation models, delivering significant gains over baseline multimodal LLMs and traditional retrieval methods in recommendation tasks.

AI researchMultimodal LLMRecommendation Systems
0 likes · 11 min read
How NoteLLM-2 Boosts Multimodal Recommendations with In-Content Learning
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Oct 16, 2024 · Artificial Intelligence

How VICTORIA Revolutionizes Multi‑Object Image Editing with Language‑Aware Diffusion

The VICTORIA algorithm, presented by Alibaba Cloud AI Platform PAI and South China University of Technology at ACM MM 2024, leverages linguistic dependency parsing to guide cross‑attention in Stable Diffusion, enabling accurate, training‑free multi‑object image editing while preserving spatial structure and achieving state‑of‑the‑art results on benchmark datasets.

AI researchStable DiffusionVICTORIA
0 likes · 10 min read
How VICTORIA Revolutionizes Multi‑Object Image Editing with Language‑Aware Diffusion
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Oct 15, 2024 · Artificial Intelligence

How VICTORIA Boosts Text‑Guided Image Editing with Language‑Aware Diffusion

The VICTORIA algorithm, presented by Alibaba Cloud's PAI team at ACM MM2024, leverages linguistic dependency parsing and cross‑attention control to overcome multi‑object editing challenges in training‑free text‑guided image editing, delivering precise, structure‑preserving results across diverse scenes.

AI researchdiffusion modelsimage manipulation
0 likes · 6 min read
How VICTORIA Boosts Text‑Guided Image Editing with Language‑Aware Diffusion
Network Intelligence Research Center (NIRC)
Network Intelligence Research Center (NIRC)
Oct 8, 2024 · Artificial Intelligence

Two NIRC Papers Accepted at NeurIPS 2024: FM-Delta Compression and GLAFF Forecasting

The Beijing University of Posts and Telecommunications' Network Intelligent Research Center (NIRC) had two papers accepted to NeurIPS 2024, presenting FM-Delta, a lossless compression technique that halves storage and cuts cloud costs by over 40%, and GLAFF, a global‑local fusion framework that markedly improves the robustness of time‑series forecasting across multiple domains.

AI researchFM-DeltaGLAFF
0 likes · 8 min read
Two NIRC Papers Accepted at NeurIPS 2024: FM-Delta Compression and GLAFF Forecasting
Fighter's World
Fighter's World
Sep 30, 2024 · Artificial Intelligence

Exploring Google NotebookLM: Use Cases, Interaction Experience, and Key Insights

The author reviews Google NotebookLM, describing how it aids deep paper reading, boosts chat willingness with guided prompts, maintains conversation coherence through self‑play insights, highlights the audio‑overview feature, and reflects on AI concepts such as the "bitter lesson" and the limits of self‑play in open scenarios.

AI researchGoogleLLM
0 likes · 22 min read
Exploring Google NotebookLM: Use Cases, Interaction Experience, and Key Insights
Kuaishou Tech
Kuaishou Tech
Sep 27, 2024 · Artificial Intelligence

XPSR: Cross‑modal Priors for Diffusion‑based Image Super‑Resolution

The paper introduces XPSR, a diffusion‑based image super‑resolution method that incorporates cross‑modal semantic priors from a large multimodal language model, achieving state‑of‑the‑art performance on both reference and no‑reference quality metrics across synthetic and real‑world video restoration tasks.

AI researchECCV2024cross‑modal priors
0 likes · 8 min read
XPSR: Cross‑modal Priors for Diffusion‑based Image Super‑Resolution
DataFunSummit
DataFunSummit
Sep 13, 2024 · Artificial Intelligence

Research on Domain Large Models by Fudan University Knowledge Workshop Lab

This article presents the Fudan University Knowledge Workshop Lab's comprehensive research on domain large models, covering background, domain adaptation, capability enhancement, collaborative workflows, challenges such as inference cost and alignment, and proposed solutions including source‑enhanced training, self‑correction mechanisms, and hybrid retrieval‑augmented generation.

AI researchDomain Adaptationknowledge graphs
0 likes · 16 min read
Research on Domain Large Models by Fudan University Knowledge Workshop Lab
Baobao Algorithm Notes
Baobao Algorithm Notes
Sep 5, 2024 · Artificial Intelligence

Why Small LLMs Are the Secret Weapon for Scaling Large Model Research

The article explains how homologous small language models—trained on the same tokenizer and data as their large counterparts—serve as cheap, fast experimental platforms that can predict large‑model performance, guide pre‑training decisions, and support techniques like distillation and reward modeling.

AI researchLLM scalingQwen2
0 likes · 13 min read
Why Small LLMs Are the Secret Weapon for Scaling Large Model Research
360 Tech Engineering
360 Tech Engineering
Aug 29, 2024 · Artificial Intelligence

FancyVideo: Towards Dynamic and Consistent Video Generation via Cross-frame Textual Guidance

FancyVideo is an open‑source UNet‑based video generation model that supports arbitrary resolutions, aspect ratios, styles, and motion dynamics by introducing a Cross‑frame Textual Guidance Module (CTGM) with temporal injectors, refiners, and boosters, achieving state‑of‑the‑art results on multiple benchmarks and enabling versatile applications such as video extension, backtracking, and frame interpolation.

AI researchUNetVideo Generation
0 likes · 6 min read
FancyVideo: Towards Dynamic and Consistent Video Generation via Cross-frame Textual Guidance
AntTech
AntTech
Aug 28, 2024 · Artificial Intelligence

Ant Group’s Selected Papers at KDD2024: Abstracts and Highlights

The article presents a curated collection of Ant Group's research papers accepted at KDD2024, summarizing each paper's title, type, link, source, relevant fields, and abstract, covering topics such as graph mining, large language models, fraud detection, recommendation systems, and multimodal medical AI.

AI researchAnt GroupKDD2024
0 likes · 31 min read
Ant Group’s Selected Papers at KDD2024: Abstracts and Highlights
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Aug 20, 2024 · Artificial Intelligence

How DAFNet Enables Efficient Sequential Editing of Large Language Models

This article introduces DAFNet, a dynamic auxiliary fusion framework that enables efficient sequential editing of large language models by injecting knowledge with reduced resource costs while preserving model reliability, generalization, and mitigating hallucination, and details its dataset, architecture, and evaluation results.

AI researchdynamic auxiliary fusionmodel editing
0 likes · 10 min read
How DAFNet Enables Efficient Sequential Editing of Large Language Models
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Aug 19, 2024 · Artificial Intelligence

How Long‑Tail Knowledge Boosts Retrieval‑Augmented Large Language Models

The paper introduces a method that classifies user queries into ordinary and long‑tail types, applying retrieval‑augmented generation only to long‑tail queries, which improves large language model efficiency and accuracy by leveraging specialized knowledge detection metrics and an extended RAG pipeline.

AI researchECE metricRetrieval-Augmented Generation
0 likes · 9 min read
How Long‑Tail Knowledge Boosts Retrieval‑Augmented Large Language Models
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Aug 11, 2024 · Artificial Intelligence

Alibaba Cloud PAI’s Breakthroughs in Chinese Diffusion, Prompting, and LLM Knowledge Editing

Recent ACL 2024 papers from Alibaba Cloud’s PAI platform showcase open‑source Chinese diffusion models, an interactive multi‑turn prompt generator, a long‑tail knowledge‑aware retrieval‑augmented LLM approach, and a dynamic fusion network for sequential model editing, all integrated into cloud services.

AI researchRetrieval-Augmented Generationdiffusion models
0 likes · 11 min read
Alibaba Cloud PAI’s Breakthroughs in Chinese Diffusion, Prompting, and LLM Knowledge Editing
21CTO
21CTO
Jul 10, 2024 · Information Security

Did a Hacker Breach OpenAI’s Internal AI Discussions? Implications for Security

A New York Times report reveals that a hacker accessed OpenAI's internal messaging system, exposing employee discussions on AI advancements and sparking concerns about foreign espionage, internal security practices, and the broader national‑security implications of AI technology.

AI researchAI securityOpenAI
0 likes · 4 min read
Did a Hacker Breach OpenAI’s Internal AI Discussions? Implications for Security
DataFunSummit
DataFunSummit
Jul 9, 2024 · Artificial Intelligence

Applying Large Language Models to Recommendation Systems at Ant Group

This article details Ant Group's research on integrating large language models into recommendation pipelines, covering background challenges, knowledge extraction, teacher‑student distillation, experimental results, and practical Q&A for improving bias, efficiency, and cold‑start performance.

AI researchAnt GroupLarge Language Models
0 likes · 14 min read
Applying Large Language Models to Recommendation Systems at Ant Group
Baobao Algorithm Notes
Baobao Algorithm Notes
Jul 9, 2024 · Artificial Intelligence

Why Step-Level DPO Is Revolutionizing LLM Math Reasoning

This article reviews recent step‑level DPO research, compares it with instance‑level DPO, explains the underlying Monte Carlo Tree Search formulation, and presents the author’s own replication experiments that demonstrate consistent performance gains across multiple LLM sizes on GSM8K and MATH benchmarks.

AI researchLLM alignmentMCTS
0 likes · 10 min read
Why Step-Level DPO Is Revolutionizing LLM Math Reasoning
Meituan Technology Team
Meituan Technology Team
Jun 27, 2024 · Artificial Intelligence

Meituan Technical Team's Three Papers Accepted at SIGIR 2024: Ad Auction Integration, Federated Recommendation, and POI Recommendation

The article highlights three Meituan research papers accepted at SIGIR 2024—covering deep automated mechanism design for ad auction, a retrieval‑enhanced vertical federated recommendation framework, and disentangled contrastive hypergraph learning for next POI recommendation—and announces an online sharing event where the authors will present their work.

AI researchAd AuctionFederated Recommendation
0 likes · 9 min read
Meituan Technical Team's Three Papers Accepted at SIGIR 2024: Ad Auction Integration, Federated Recommendation, and POI Recommendation
Alibaba Cloud Developer
Alibaba Cloud Developer
Jun 27, 2024 · Artificial Intelligence

How to Supercharge Retrieval‑Augmented Generation: Papers, Techniques, and Real‑World Tips

This article surveys the main challenges of deploying large language models, introduces key RAG optimization papers such as RAPTOR, Self‑RAG, and CRAG, and compiles practical engineering tricks—including chunking, query rewriting, hybrid and progressive retrieval—to help practitioners build more accurate and efficient RAG systems.

AI researchLLM OptimizationRAG
0 likes · 22 min read
How to Supercharge Retrieval‑Augmented Generation: Papers, Techniques, and Real‑World Tips
Xiaohongshu Tech REDtech
Xiaohongshu Tech REDtech
Jun 20, 2024 · Artificial Intelligence

Xiaohongshu 2024 Large Model Frontier Paper Sharing Live Event

On June 27, 2024, Xiaohongshu’s technical team will livestream a two‑hour session across WeChat Channels, Bilibili, Douyin and Xiaohongshu, showcasing six top‑conference papers on large‑model advances—including early‑stopping and fine‑grained self‑consistency, novel evaluation methods, negative‑sample‑assisted distillation, and LLM‑based note recommendation—followed by a Q&A and recruitment briefing.

AI researchLarge Language ModelsModel Evaluation
0 likes · 12 min read
Xiaohongshu 2024 Large Model Frontier Paper Sharing Live Event
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Jun 18, 2024 · Artificial Intelligence

Free-Prompt-Editing: Efficient Text-Guided Image Editing with Stable Diffusion

The paper introduces Free-Prompt-Editing (FPE), a novel, efficient algorithm for text‑guided image editing that leverages probe analysis of cross‑ and self‑attention maps in Stable Diffusion, demonstrates its superiority over existing methods through extensive experiments, and provides open‑source implementation for both synthetic and real‑image editing.

AI researchStable Diffusionattention maps
0 likes · 12 min read
Free-Prompt-Editing: Efficient Text-Guided Image Editing with Stable Diffusion
Alibaba Cloud Big Data AI Platform
Alibaba Cloud Big Data AI Platform
Jun 17, 2024 · Artificial Intelligence

How Free-Prompt-Editing Revolutionizes Text-Guided Image Editing with Stable Diffusion

The paper introduces Free-Prompt-Editing, a concise and efficient algorithm that replaces self‑attention maps during denoising to achieve high‑quality text‑guided image edits without source prompts, and demonstrates its superiority over existing methods on both synthetic and real images.

AI researchAttention MechanismsFree-Prompt-Editing
0 likes · 6 min read
How Free-Prompt-Editing Revolutionizes Text-Guided Image Editing with Stable Diffusion
DataFunTalk
DataFunTalk
Jun 15, 2024 · Artificial Intelligence

Research on Domain Large Models by Fudan University Knowledge Factory Lab

This article presents Fudan University's Knowledge Factory Lab research on domain large models, covering background, challenges, data selection, source‑enhanced tagging, capability improvements, self‑correction, collaborative workflows, and retrieval‑augmented generation for practical AI deployment.

AI researchDomain AdaptationKnowledge Graph
0 likes · 16 min read
Research on Domain Large Models by Fudan University Knowledge Factory Lab
DataFunSummit
DataFunSummit
Jun 6, 2024 · Artificial Intelligence

MetaGPT: Multi‑Agent Collaboration and Agent Capability Enhancement

This article introduces MetaGPT, an open‑source multi‑agent framework that leverages large language models to automate software development, data science, and simulation tasks, detailing its development, impact, experimental results, memory and reasoning enhancements, and comparisons with related systems.

AI researchAgent MemoryLLM agents
0 likes · 21 min read
MetaGPT: Multi‑Agent Collaboration and Agent Capability Enhancement
NewBeeNLP
NewBeeNLP
May 28, 2024 · Artificial Intelligence

How Generative Models Are Redefining Recommendation Systems

This article reviews recent advances in generative recommendation, highlighting challenges such as item representation and multimodal fusion, and summarizing four key research papers that propose novel tokenization, collaborative integration, and transformer-based multimodal approaches to improve recommendation performance.

AI researchCollaborative FilteringGenerative Recommendation
0 likes · 8 min read
How Generative Models Are Redefining Recommendation Systems
360 Tech Engineering
360 Tech Engineering
May 17, 2024 · Artificial Intelligence

360VL: An Open‑Source Multimodal Large Language Model Based on Llama‑3‑70B

The article introduces 360VL, an open‑source multimodal large language model built on Llama‑3‑70B, describes its novel C‑abs bridge architecture for high‑resolution visual understanding, outlines the two‑stage training with bilingual data, and presents benchmark results showing superior performance over prior LMMs.

AI researchLlama3Multimodal
0 likes · 8 min read
360VL: An Open‑Source Multimodal Large Language Model Based on Llama‑3‑70B
NewBeeNLP
NewBeeNLP
May 15, 2024 · Artificial Intelligence

How Large Language Models and Knowledge Graphs Can Boost Each Other

This talk reviews recent advances in large language models, compares them with knowledge graphs, explores how LLMs enhance knowledge extraction and completion, examines how knowledge graphs aid LLM evaluation and safe deployment, and outlines future interactive integration between the two technologies.

AI researchLarge Language ModelsModel Evaluation
0 likes · 13 min read
How Large Language Models and Knowledge Graphs Can Boost Each Other
Rare Earth Juejin Tech Community
Rare Earth Juejin Tech Community
May 15, 2024 · Artificial Intelligence

OpenAI Unveils GPT‑4o: An Omni‑Capable Multimodal Model Offered Free to All Users

OpenAI introduced GPT‑4o, a free, omni‑capable multimodal model that processes text, audio, and images together, delivers near‑human response latency, showcases impressive live demos, and will soon be available via a discounted API, marking a significant step forward in end‑to‑end AI research.

AI researchGPT-4oOpenAI
0 likes · 7 min read
OpenAI Unveils GPT‑4o: An Omni‑Capable Multimodal Model Offered Free to All Users
21CTO
21CTO
Apr 8, 2024 · Artificial Intelligence

How Naver’s HyperCLOVA X Advances Multilingual AI for Asian Languages

Naver’s newly unveiled HyperCLOVA X large‑language model, detailed in an arXiv technical report, claims superior cross‑lingual reasoning for Asian languages, especially Korean, by pre‑training on a data mix of Korean, multilingual text and code, achieving state‑of‑the‑art translation and multilingual capabilities.

AI researchHyperCLOVA XKorean NLP
0 likes · 4 min read
How Naver’s HyperCLOVA X Advances Multilingual AI for Asian Languages
DaTaobao Tech
DaTaobao Tech
Mar 29, 2024 · Artificial Intelligence

Text-to-SQL with Large Language Models: DIN-SQL Approach

The DIN‑SQL approach enhances Text‑to‑SQL performance by using large language models in a decomposed in‑context learning framework with schema linking, query classification, SQL generation, and self‑correction modules, achieving state‑of‑the‑art 85.3% execution accuracy on the Spider benchmark by breaking complex queries into manageable sub‑tasks.

AI researchData AnalysisDatabase Querying
0 likes · 34 min read
Text-to-SQL with Large Language Models: DIN-SQL Approach
Open Source Tech Hub
Open Source Tech Hub
Mar 17, 2024 · Artificial Intelligence

What Is Grok? Inside Elon Musk’s New Open‑Source LLM and the ‘Grokking’ Phenomenon

Elon Musk announced the open‑source release of Grok, xAI’s new large‑language‑model chatbot, while recalling his lawsuit against OpenAI; the article explains Grok’s rapid development, links to the GitHub repository, summarizes the seminal “Grokking” research paper that describes a sudden generalization breakthrough in neural networks, and provides reference links.

AI researchGrokGrokking
0 likes · 3 min read
What Is Grok? Inside Elon Musk’s New Open‑Source LLM and the ‘Grokking’ Phenomenon
JD Retail Technology
JD Retail Technology
Mar 12, 2024 · Artificial Intelligence

Multimodal Large Models: Recent Advances, Industry Impact, and Challenges – An Expert Interview

In a detailed interview, Tsinghua researcher Zhao Sicheng and JD Retail senior director Peng Changping discuss the latest progress in multimodal large models, their practical applications in advertising and e‑commerce, persistent challenges such as hallucinations and data alignment, and the skills engineers need to thrive in the emerging AI era.

AI researche-commercelarge models
0 likes · 19 min read
Multimodal Large Models: Recent Advances, Industry Impact, and Challenges – An Expert Interview
Baobao Algorithm Notes
Baobao Algorithm Notes
Mar 10, 2024 · Artificial Intelligence

Unlocking Large Model Power: 5 Effective Model Fusion Techniques Explained

This article examines why ensemble methods are crucial for large language models, outlines five core fusion strategies—including model integration, probability integration, graft learning, crowdsourced voting, and Mixture of Experts—provides implementation details, pseudo‑code, and discusses practical challenges and recent research advances.

AI researchMixture of Expertsensemble methods
0 likes · 16 min read
Unlocking Large Model Power: 5 Effective Model Fusion Techniques Explained
DataFunTalk
DataFunTalk
Mar 10, 2024 · Artificial Intelligence

Aligning Graph Models with Large Language Models for Open-Task Scenarios

This talk presents GraphTranslator, a framework that bridges pretrained graph models and large language models to enable unified handling of both predefined and open-ended graph analysis tasks by translating node representations into language tokens and training an alignment producer for node‑text pairs.

AI researchLarge Language ModelsModel Alignment
0 likes · 3 min read
Aligning Graph Models with Large Language Models for Open-Task Scenarios
Sohu Tech Products
Sohu Tech Products
Mar 6, 2024 · Artificial Intelligence

Analysis of OpenAI Sora: Data Engineering, Network Architecture, and World Model Implications

OpenAI’s Sora video model unifies image and video data into latent spacetime patches via a VAE, trains on original resolutions with GPT‑4‑expanded captions, employs a Diffusion Transformer backbone for patch‑wise denoising, and demonstrates 3D‑consistent, long‑term world‑model capabilities that hint at a unified computer‑vision paradigm and steps toward AGI.

AI researchOpenAI SoraTransformer
0 likes · 9 min read
Analysis of OpenAI Sora: Data Engineering, Network Architecture, and World Model Implications
NetEase Smart Enterprise Tech+
NetEase Smart Enterprise Tech+
Feb 28, 2024 · Artificial Intelligence

Mastering Multi-Task Learning: Network Designs & Loss Balancing

This article reviews the challenges of multi‑task learning, compares various network architectures such as hard‑parameter sharing, MMoE, CGC, and PLE, and examines loss‑balancing techniques like GradNorm, Dynamic Weight Average and task‑prioritization, offering insights on how to mitigate the “seesaw” effect and improve overall performance.

AI researchdynamic weightinggradient normalization
0 likes · 15 min read
Mastering Multi-Task Learning: Network Designs & Loss Balancing
21CTO
21CTO
Feb 17, 2024 · Artificial Intelligence

How OpenAI’s Sora Is Pushing Video Generation to New Frontiers

OpenAI’s Sora model demonstrates large‑scale text‑conditional video generation using a diffusion transformer that operates on spatiotemporal patches, supporting variable durations, resolutions, and aspect ratios while showcasing emergent simulation abilities, flexible sampling, and multimodal editing capabilities, though it still has notable limitations.

AI researchMultimodalSora
0 likes · 19 min read
How OpenAI’s Sora Is Pushing Video Generation to New Frontiers
NewBeeNLP
NewBeeNLP
Feb 12, 2024 · Artificial Intelligence

Beyond Dual‑Tower: Advanced Distillation and Interaction Techniques for Recommendation Systems

This article reviews recent advances that enhance dual‑tower recommendation models by injecting interaction information through various knowledge‑distillation strategies and interaction‑enhanced architectures, summarizing methods such as PFD, ENDX, TRMD, VIRT, Distilled‑DualEncoder, ERNIE‑Search, ColBert, IntTower and MVKE.

AI researchdual-towerinteraction modeling
0 likes · 13 min read
Beyond Dual‑Tower: Advanced Distillation and Interaction Techniques for Recommendation Systems
IT Services Circle
IT Services Circle
Jan 3, 2024 · Artificial Intelligence

Sergey Brin’s Role in Google’s Gemini AI Model and His Return to Technical Work

The article recounts Sergey Brin’s surprising appearance as a core contributor to Google’s Gemini AI model, tracing his early technical career, his semi‑retirement from Alphabet, the company’s AI challenges after ChatGPT, and how Brin returned to help develop Gemini, highlighting internal reactions and his lasting influence.

AI researchArtificial IntelligenceGemini
0 likes · 9 min read
Sergey Brin’s Role in Google’s Gemini AI Model and His Return to Technical Work
DataFunTalk
DataFunTalk
Dec 21, 2023 · Artificial Intelligence

Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning – Best Long Paper at EMNLP 2023

At EMNLP 2023, the joint WeChat AI and Peking University paper 'Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning' won the Best Long Paper award, revealing that label tokens act as anchors driving information aggregation in shallow layers and prediction flow in deep layers, and proposing methods to improve and diagnose in‑context learning.

AI researchInformation FlowLarge Language Models
0 likes · 13 min read
Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning – Best Long Paper at EMNLP 2023
Baobao Algorithm Notes
Baobao Algorithm Notes
Nov 9, 2023 · Artificial Intelligence

Building High‑Performance Vertical Domain LLMs: From Continued Pre‑Training to Retrieval‑Augmented Generation

This article systematically explains how to create vertical domain large language models by continuing pre‑training on domain data, constructing fine‑tuning datasets with self‑instruct, reducing hallucinations, and integrating knowledge retrieval, while also reviewing related papers, products, and system architectures.

AI researchKnowledge Retrievalself-instruct
0 likes · 21 min read
Building High‑Performance Vertical Domain LLMs: From Continued Pre‑Training to Retrieval‑Augmented Generation
Ximalaya Technology Team
Ximalaya Technology Team
Oct 10, 2023 · Artificial Intelligence

MiniGPT-5: A Novel Multimodal Generation Model for Coherent Text-Image Synthesis

MiniGPT-5 is a novel multimodal generation model using generative vokens to interleave text and image synthesis, integrating Stable Diffusion and LLMs with a two-stage training that requires no domain-specific annotations, achieving state‑of‑the‑art coherence and quality on benchmarks like CC3M, VIST, and MMDialog.

AI researchStable DiffusionVision Transformer
0 likes · 9 min read
MiniGPT-5: A Novel Multimodal Generation Model for Coherent Text-Image Synthesis
Baobao Algorithm Notes
Baobao Algorithm Notes
Oct 8, 2023 · Interview Experience

Must‑Know Large‑Model Interview Questions for RLHF Candidates

The article shares a practitioner’s transition story from reinforcement‑learning‑focused game AI to large‑model work, outlines the challenges faced during job hunting at major Chinese tech firms, and provides a curated list of 23 technical interview questions covering PPO, RLHF, dataset evaluation, model fine‑tuning, and broader LLM concepts.

AI researchInterview PreparationLLM
0 likes · 10 min read
Must‑Know Large‑Model Interview Questions for RLHF Candidates
DataFunTalk
DataFunTalk
Sep 26, 2023 · Artificial Intelligence

MiniGPT-4: Enhancing Vision‑Language Understanding with Large Language Models

This article presents MiniGPT-4, a multimodal system that combines a frozen visual encoder (Q‑Former + ViT) with an open‑source large language model (Vicuna), describes its motivation, training pipeline, demo capabilities, observed limitations, and includes a brief Q&A session.

AI researchImage CaptioningMiniGPT-4
0 likes · 15 min read
MiniGPT-4: Enhancing Vision‑Language Understanding with Large Language Models
AntTech
AntTech
Sep 15, 2023 · Artificial Intelligence

Ant Group Unveils Large Graph Model (LGM) Merging Graph Computing with Large Language Models

At the 2023 Bund Conference, Ant Group presented the Large Graph Model (LGM), a research effort that combines graph computing, graph learning, and large language models to enrich heterogeneous graph data and enable more precise insights for complex digital applications, with results accepted at WWW 2023.

AI researchAnt GroupLarge Graph Model
0 likes · 6 min read
Ant Group Unveils Large Graph Model (LGM) Merging Graph Computing with Large Language Models
Baobao Algorithm Notes
Baobao Algorithm Notes
Aug 18, 2023 · Artificial Intelligence

Unlocking Domain-Specific Large Model Training: Proven Tricks and Pitfalls

This article shares practical techniques for domain‑specific large model continue pre‑training, including data selection, mixing ratios with general data, multi‑task instruction pre‑training, resource‑aware fine‑tuning strategies, evaluation set design, vocabulary considerations, and deployment constraints for 7‑13B models.

AI researchModel EvaluationSFT
0 likes · 9 min read
Unlocking Domain-Specific Large Model Training: Proven Tricks and Pitfalls
21CTO
21CTO
Aug 15, 2023 · Artificial Intelligence

Why Do Neural Networks Suddenly ‘Grok’ After Long Training? Insights from Google

Google’s recent research reveals that when small neural networks are trained for extended periods on tasks like modular addition, they can abruptly shift from memorizing training data to genuinely generalizing—a sudden “grokking” phenomenon driven by weight decay and the emergence of periodic weight structures.

AI researchGrokkingMLP
0 likes · 9 min read
Why Do Neural Networks Suddenly ‘Grok’ After Long Training? Insights from Google
Rare Earth Juejin Tech Community
Rare Earth Juejin Tech Community
Aug 1, 2023 · Artificial Intelligence

Do Language Models Learn Language in the Same Stages as Children? An Analysis of GPT‑2 Developmental Trajectories

This article reviews a study that compares the stage‑wise language acquisition of infants with the learning trajectory of GPT‑2, using linguistic probes and statistical tests to determine whether deep language models follow sequential or parallel learning patterns similar to children.

AI researchGPT-2developmental learning
0 likes · 17 min read
Do Language Models Learn Language in the Same Stages as Children? An Analysis of GPT‑2 Developmental Trajectories
Baidu Geek Talk
Baidu Geek Talk
Jul 26, 2023 · Artificial Intelligence

Insights on AIGC Development and Commercial Applications by Baidu's Chief Architect

Baidu’s chief architect Li Shuanglong outlined how AIGC, driven by advanced large‑language and multimodal models, is already powering commercial tools such as automated copywriting, 2D digital‑human video creation and lead‑generation chatbots, while emphasizing future progress in engineering scalability, algorithmic fidelity, data quality, and scenario‑focused applications.

AI commercializationAI researchAIGC
0 likes · 8 min read
Insights on AIGC Development and Commercial Applications by Baidu's Chief Architect
DataFunSummit
DataFunSummit
Jun 28, 2023 · Artificial Intelligence

OPPO's CHAOS Pretrained Large Model and GammaE Knowledge‑Graph Multi‑hop Reasoning: Techniques and Insights

This article presents OPPO Research Institute's recent advances in large‑model AI, detailing the CHAOS pretrained model that topped the CLUE leaderboard, the knowledge‑enhanced training pipeline, and the GammaE model for multi‑hop reasoning over knowledge graphs, together with experimental results and practical training tips.

AI researchGammaEKnowledge Graph
0 likes · 20 min read
OPPO's CHAOS Pretrained Large Model and GammaE Knowledge‑Graph Multi‑hop Reasoning: Techniques and Insights
DataFunSummit
DataFunSummit
May 31, 2023 · Artificial Intelligence

Evolution of Face Detection Techniques: Datasets, Research Directions, and Future Work

This article reviews the evolution of face detection, covering the Widely‑Face dataset, major research directions such as feature fusion, label assignment, auxiliary supervision, anchor‑free methods, NAS‑based designs, summarizes key papers from S3FD to MogFace, introduces ModelScope implementations, and outlines future challenges and opportunities.

AI researchModel Evaluationcomputer vision
0 likes · 13 min read
Evolution of Face Detection Techniques: Datasets, Research Directions, and Future Work
Kuaishou Tech
Kuaishou Tech
Apr 28, 2023 · Artificial Intelligence

How Hyper‑Actor Critic Redefines Reinforcement Learning for Recommendation Systems

This article presents the Hyper‑Actor Critic (HAC) framework that splits reinforcement‑learning policies into continuous hyper‑actions and effective recommendation lists, introduces alignment and supervised losses, and demonstrates superior performance on an online simulator compared to existing RL and supervised methods.

AI researchRecommendation Systemshyper-actor critic
0 likes · 9 min read
How Hyper‑Actor Critic Redefines Reinforcement Learning for Recommendation Systems
Architect
Architect
Apr 27, 2023 · Artificial Intelligence

Survey of Large Language Model Research: From GPT‑1 to ChatGPT and Open‑Source Alternatives

This article provides a comprehensive overview of the development of large language models, reviewing classic papers from GPT‑1 through GPT‑4, discussing open‑source implementations such as LLaMA, Alpaca, GLM, and ChatGLM, and analyzing training methods, datasets, and future research directions.

AI researchGPTLarge Language Models
0 likes · 36 min read
Survey of Large Language Model Research: From GPT‑1 to ChatGPT and Open‑Source Alternatives
21CTO
21CTO
Apr 21, 2023 · Artificial Intelligence

Essential AI Reading List: LLMs, AutoGPT, Distributed Training & More

This curated collection highlights the latest open‑source LLM breakthroughs, comprehensive surveys, AutoGPT developments, distributed training pitfalls, and practical tools for AI engineers, providing concise descriptions and direct links to each resource for deeper exploration.

AI researchAutoGPTDistributed Training
0 likes · 10 min read
Essential AI Reading List: LLMs, AutoGPT, Distributed Training & More
IT Architects Alliance
IT Architects Alliance
Apr 20, 2023 · Artificial Intelligence

Overview of Prominent Large Language Models and Instruction‑Finetuned Variants

This article provides a comprehensive overview of major large language models—including GPT series, T5, LaMDA, LLaMA, BLOOM, and others—detailing their architectures, parameter scales, open‑source status, and the evolution of instruction‑fine‑tuning techniques that improve zero‑shot and few‑shot performance.

AI researchInstruction TuningLLM comparison
0 likes · 24 min read
Overview of Prominent Large Language Models and Instruction‑Finetuned Variants
Architect
Architect
Apr 19, 2023 · Artificial Intelligence

Emergence in Large Language Models: Phenomena, Explanations, and Implications

This article reviews the emergence phenomena observed in large language models, explains how model scale, in‑context learning and chain‑of‑thought prompting contribute to sudden performance gains, discusses small‑model alternatives, and explores the relationship between emergence and the training‑time Grokking effect.

AI researchGrokkingLarge Language Models
0 likes · 13 min read
Emergence in Large Language Models: Phenomena, Explanations, and Implications
DataFunTalk
DataFunTalk
Apr 19, 2023 · Artificial Intelligence

Is the Daily Emergence of Large Language Models Beneficial?

The article examines the rapid proliferation of large language models, weighing both the opportunities for experimentation and the drawbacks of noise, and argues that establishing authoritative Chinese LLM evaluation benchmarks is essential to guide meaningful progress in the field.

AI researchLLM evaluationLarge Language Models
0 likes · 7 min read
Is the Daily Emergence of Large Language Models Beneficial?
21CTO
21CTO
Apr 9, 2023 · Artificial Intelligence

8 Open-Source ChatGPT Alternatives You Can Deploy Today

This article surveys eight popular open‑source ChatGPT alternatives, detailing each model’s size, training data, performance relative to proprietary systems, and providing links to code repositories, demos, and papers for developers interested in building or researching large language models.

AI researchChatGPT alternativesmodel comparison
0 likes · 8 min read
8 Open-Source ChatGPT Alternatives You Can Deploy Today
21CTO
21CTO
Mar 31, 2023 · Artificial Intelligence

From Student to AI Pioneer: Ilya Sutskever’s Journey Behind ChatGPT

This article chronicles Ilya Sutskever’s two‑decade rise from a young researcher to a leading figure in artificial intelligence, highlighting his early mentorship, breakthroughs in image recognition, language translation, the founding of OpenAI, and the development of GPT and DALL‑E models.

AI researchGPTIlya Sutskever
0 likes · 13 min read
From Student to AI Pioneer: Ilya Sutskever’s Journey Behind ChatGPT
DataFunSummit
DataFunSummit
Mar 21, 2023 · Artificial Intelligence

Interview with Huawei Noah's Ark Lab Senior Researcher Zhou Min on Graph Machine Learning: Research, Deployment, Challenges, and Trends

In this DataFun interview, Huawei Noah's Ark Lab senior researcher Zhou Min discusses the state of graph machine learning in academia and industry, covering algorithmic foundations, model variants, practical applications, scalability challenges, and future directions for more universal feature extraction across domains.

AI researchDataFunGraph Machine Learning
0 likes · 9 min read
Interview with Huawei Noah's Ark Lab Senior Researcher Zhou Min on Graph Machine Learning: Research, Deployment, Challenges, and Trends
DataFunTalk
DataFunTalk
Mar 16, 2023 · Artificial Intelligence

Technical Optimizations and Breakthroughs of GPT‑4: Multimodal Capabilities, Alignment Strategies, and Predictable Scaling

The article summarizes the technical innovations behind GPT‑4, highlighting its multimodal abilities, improved alignment methods, scaling‑law‑based performance prediction, and remaining limitations, while referencing the official OpenAI technical report and community analyses.

AI researchAlignmentGPT-4
0 likes · 10 min read
Technical Optimizations and Breakthroughs of GPT‑4: Multimodal Capabilities, Alignment Strategies, and Predictable Scaling
Python Programming Learning Circle
Python Programming Learning Circle
Mar 10, 2023 · Artificial Intelligence

Google's i‑S2R and GoalsEye: Robot Table‑Tennis Learning from Human Interaction

The article explains how Google's i‑S2R and GoalsEye projects use iterative simulation‑to‑real training, behavior cloning and goal‑conditioned learning to enable robots to play table‑tennis with humans, highlighting the challenges, experimental setup, and performance improvements achieved across player skill levels.

AI researchSim2Realbehavior cloning
0 likes · 6 min read
Google's i‑S2R and GoalsEye: Robot Table‑Tennis Learning from Human Interaction
21CTO
21CTO
Feb 27, 2023 · Artificial Intelligence

What’s Next for Large Language Models? Emerging Trends Shaping AI

The article explores three emerging directions for next‑generation large language models—self‑generated training data, built‑in verification with external retrieval, and massive sparse‑expert architectures—highlighting recent research, practical challenges, and their potential to reshape AI development.

AI researchLarge Language Modelsgenerative AI
0 likes · 17 min read
What’s Next for Large Language Models? Emerging Trends Shaping AI
DataFunTalk
DataFunTalk
Feb 15, 2023 · Artificial Intelligence

Three Emerging Directions for Next‑Generation Large Language Models

The article outlines three promising research avenues—self‑generated training data, model‑driven fact‑checking, and sparse expert architectures—that could shape the next wave of large language model innovation and address current limitations such as data scarcity and hallucinations.

AI researchLarge Language Modelsmodel self‑improvement
0 likes · 14 min read
Three Emerging Directions for Next‑Generation Large Language Models
Architect
Architect
Feb 9, 2023 · Artificial Intelligence

Emergent Abilities of Large Language Models: Complex Reasoning, Knowledge Reasoning, and Out‑of‑Distribution Robustness

This article reviews recent research on the emergent abilities of large language models—such as chain‑of‑thought reasoning, knowledge retrieval without external sources, and robustness to distribution shifts—examining scaling laws, model size thresholds, and the open questions surrounding a potential paradigm shift from fine‑tuning to in‑context learning.

AI researchLarge Language Modelschain-of-thought prompting
0 likes · 23 min read
Emergent Abilities of Large Language Models: Complex Reasoning, Knowledge Reasoning, and Out‑of‑Distribution Robustness
DataFunSummit
DataFunSummit
Feb 7, 2023 · Artificial Intelligence

How to Evaluate OpenAI's Super Conversational Model ChatGPT?

This article compiles three highly upvoted Zhihu answers that examine OpenAI's ChatGPT, discussing its breakthrough impact on NLP, visual in‑context learning, reinforcement‑learning‑from‑human‑feedback, and the broader implications for AI research and development.

AI researchChatGPTLarge Language Models
0 likes · 10 min read
How to Evaluate OpenAI's Super Conversational Model ChatGPT?
21CTO
21CTO
Jan 13, 2023 · Artificial Intelligence

How Google’s Muse Is Redefining Text‑to‑Image Generation with Parallel Decoding

Google’s new Muse model, a Transformer‑based text‑to‑image system running on TPUv4, claims to generate 256×256 images in 0.5 seconds—far faster than Imagen—while delivering unprecedented photorealism and deep language understanding through parallel decoding and large‑scale LLM‑conditioned training.

AI researchGoogle MuseLLM conditioning
0 likes · 4 min read
How Google’s Muse Is Redefining Text‑to‑Image Generation with Parallel Decoding
DataFunTalk
DataFunTalk
Jan 10, 2023 · Artificial Intelligence

Paradigm Shifts in Large Language Model Research and Future Directions

The article reviews the evolution of large language models from the pre‑GPT‑3 era to the present, analyzes the conceptual and technical gaps between Chinese and global research, and outlines key future research directions such as scaling laws, prompting techniques, multimodal training, and efficient model architectures.

AI researchChatGPTLLM
0 likes · 73 min read
Paradigm Shifts in Large Language Model Research and Future Directions
Xiaohongshu Tech REDtech
Xiaohongshu Tech REDtech
Jan 3, 2023 · Artificial Intelligence

Insights into ChatGPT: Capabilities, Limitations, and Implications for AI Research

During Xiaohongshu’s REDtech livestream, AI researchers examined ChatGPT’s rapid adoption, versatile task performance, and underlying large‑scale pre‑training with in‑context learning, while highlighting persistent hallucinations, weak reasoning, high costs, and limited search‑engine replacement potential, and emphasized the importance of RLHF‑driven human feedback for future multimodal AI research.

AI researchChatGPTLarge Language Models
0 likes · 14 min read
Insights into ChatGPT: Capabilities, Limitations, and Implications for AI Research
Meituan Technology Team
Meituan Technology Team
Nov 17, 2022 · Artificial Intelligence

Overview of Recent Meituan Visual Intelligence Research Papers on Content Production, Distribution, and Model Quantization

Meituan’s Visual Intelligence team recently published eight top‑conference papers that advance weakly supervised segmentation, future‑aware captioning, panoptic narrative grounding, video‑text retrieval, open‑vocabulary detection, counterfactual image‑text matching, zero‑shot video classification, and efficient Vision‑Transformer quantization, all directly boosting real‑world content creation, distribution, and model efficiency.

AI researchImage CaptioningModel Quantization
0 likes · 19 min read
Overview of Recent Meituan Visual Intelligence Research Papers on Content Production, Distribution, and Model Quantization
AntTech
AntTech
Sep 27, 2022 · Artificial Intelligence

Ant Group’s Research Institute Publishes Four NeurIPS 2022 Papers on Advanced Computer Vision and AI

Ant Group’s Ant Technology Research Institute had four papers from its Visual Intelligence Lab accepted at NeurIPS 2022, covering rank diminishing in deep networks, geometry‑aware 3D image synthesis, dynamic discriminators for GANs, and uncertainty‑aware hierarchical refinement for incremental classification, highlighting the institute’s cutting‑edge AI research.

AI researchGANsNeurIPS
0 likes · 8 min read
Ant Group’s Research Institute Publishes Four NeurIPS 2022 Papers on Advanced Computer Vision and AI
DataFunTalk
DataFunTalk
Nov 1, 2021 · Artificial Intelligence

Reflections on Working as an Algorithm Engineer at Meituan and the Rise of Contrastive Learning

The author shares personal experiences as a Meituan algorithm engineer, emphasizing the critical role of labeled data, the emergence of contrastive (self‑supervised) learning across computer vision, NLP, and recommendation systems, and offers practical advice for algorithm engineers to stay competitive.

AI researchAlgorithm EngineeringMeituan
0 likes · 8 min read
Reflections on Working as an Algorithm Engineer at Meituan and the Rise of Contrastive Learning
Kuaishou Tech
Kuaishou Tech
Oct 28, 2021 · Artificial Intelligence

Kuaishou Showcases Multimedia Research at ACM MM2021 and Announces Strategic Collaboration with CCF‑MM Committee

At ACM MM2021 in Chengdu, Kuaishou presented two accepted papers on recommendation and outfit compatibility, won the Grand Challenge with its DAP congestion‑control system, and forged a strategic partnership with the CCF‑MM committee to deepen multimedia research collaboration across academia and industry.

ACM MM2021AI researchKuaishou
0 likes · 8 min read
Kuaishou Showcases Multimedia Research at ACM MM2021 and Announces Strategic Collaboration with CCF‑MM Committee
Youku Technology
Youku Technology
Sep 29, 2021 · Artificial Intelligence

Reducing the Covariate Shift by Mirror Samples in Cross Domain Alignment

By constructing virtual mirror samples that occupy identical positions across source and target domains, the authors eliminate covariate shift while preserving distribution structure, enabling superior unsupervised domain adaptation that achieves state‑of‑the‑art performance on Office and VisDA benchmarks and improves real‑world lighting and gender‑recognition tasks.

AI researchDomain AdaptationSOTA
0 likes · 3 min read
Reducing the Covariate Shift by Mirror Samples in Cross Domain Alignment
DataFunTalk
DataFunTalk
Jul 1, 2021 · Artificial Intelligence

Pre‑Trained Models: Past, Present, and Future – A Comprehensive Survey

This article surveys the evolution of pre‑trained models, covering the origins of transfer and self‑supervised learning, the rise of transformer‑based PTMs such as BERT and GPT, efficient architecture designs, multimodal and multilingual extensions, theoretical analyses, and future research directions for scalable and robust AI systems.

AI researchEfficient TrainingLarge Language Models
0 likes · 27 min read
Pre‑Trained Models: Past, Present, and Future – A Comprehensive Survey
AntTech
AntTech
Mar 3, 2021 · Artificial Intelligence

Ant Group Intelligent Service Research Overview: NLP, Dialogue, Recommendation, and Anti‑fraud Papers

The article presents a comprehensive overview of Ant Group's intelligent service research, summarizing recent AI‑focused papers on text classification, stance detection, data augmentation, knowledge distillation for ranking, reinforcement‑learning‑based dialogue clarification, behavior‑cloning dialogue systems, anti‑fraud outbound bots, tag‑based service recommendation, and multi‑agent service groups, while also highlighting future directions and recruitment opportunities.

AI researchAnti‑fraudData Augmentation
0 likes · 17 min read
Ant Group Intelligent Service Research Overview: NLP, Dialogue, Recommendation, and Anti‑fraud Papers
JD Cloud Developers
JD Cloud Developers
Feb 10, 2021 · Artificial Intelligence

Three JD Tech AI Papers Shine at ICASSP 2021

At ICASSP 2021, JD Tech presented three AI research papers—introducing a Neural Kalman Filtering framework for speech enhancement, a cross‑utterance BERT‑based prosody modeling method for end‑to‑end speech synthesis, and a self‑supervised conversational query rewriting approach—each demonstrating superior performance over existing baselines on benchmark datasets.

AI researchICASSP 2021Self-supervised Learning
0 likes · 9 min read
Three JD Tech AI Papers Shine at ICASSP 2021
JD Cloud Developers
JD Cloud Developers
Nov 2, 2020 · Artificial Intelligence

This Week’s Tech Highlights: AI Research Breakthroughs, 5G Surge, Multi‑Cloud DB & More

The newsletter recaps recent tech developments, including JD's four AI papers at Interspeech 2020, Shenzhen's supercomputing boost, T‑Mobile's mid‑band 5G expansion, Apple's upcoming A14T iMac processor, MongoDB Atlas multi‑cloud support, Wikimedia's migration to GitLab, and advances in graph neural network pre‑training and deep clustering.

5G expansionAI researchApple Silicon
0 likes · 9 min read
This Week’s Tech Highlights: AI Research Breakthroughs, 5G Surge, Multi‑Cloud DB & More
21CTO
21CTO
Jan 31, 2020 · Artificial Intelligence

How Microsoft’s First Chinese AI Fellow Is Driving Speech and Language Breakthroughs

Microsoft appointed its first Chinese Global Technical Fellow, Huang Xuedong, as the company’s Global AI CTO, overseeing Azure’s speech, translation, vision, and language services, while highlighting his groundbreaking achievements such as achieving human‑level word error rates and leading AI research teams.

AI researchArtificial IntelligenceAzure
0 likes · 7 min read
How Microsoft’s First Chinese AI Fellow Is Driving Speech and Language Breakthroughs
Hulu Beijing
Hulu Beijing
Apr 2, 2019 · Artificial Intelligence

From Object Detection to Language Models: A Deep Dive into AI Advances

This article surveys the evolution of object detection models—comparing one‑stage and two‑stage approaches, their performance trade‑offs, and recent state‑of‑the‑art methods—while also outlining key concepts and breakthroughs in natural language processing, highlighting the impact of deep‑learning models such as BERT.

AI researchBERTdeep learning
0 likes · 14 min read
From Object Detection to Language Models: A Deep Dive into AI Advances
ITPUB
ITPUB
Feb 23, 2019 · Artificial Intelligence

Explore a 1.59 Million Image NSFW Dataset with 159 Fine-Grained Categories

A data scientist from Besedo has open‑sourced a massive NSFW image dataset containing 1.589 million pictures, organized into 159 primary categories and further sub‑categories, with download scripts and GitHub links, requiring about 500 GB of storage and cautioning against viewing in the office.

AI researchGitHubLarge Dataset
0 likes · 3 min read
Explore a 1.59 Million Image NSFW Dataset with 159 Fine-Grained Categories