SenseTime Announces Official Launch of Large Language Model Application 'SenseChat'


According to a report on July 24th, Tencent announced on July 23rd that it would merge the Yuan Multimodal Model Department with the Large Language Model Department to form the Basic Model Department, led by Chief AI Scientist YAO Shunyu. This move aims to enhance R&D and collaboration efficiency, and strive to achieve the intelligence upper limit of full modal models. There were hints of this integration earlier, as YAO Shunyu had already taken charge of the Large Language Model team in December last year. Now, combining both sides means Tencent is concentrating resources to promote the deep integration of multimodal and language models, accelerating the construction of a new generation of unified basic models, and moving towards higher peaks of full modal intelligence.
Epoch AI research shows that mainstream AI text detectors can nearly perfectly identify ordinary AI-generated content, but when large language models deliberately mimic the writing styles of specific authors, their accuracy drops significantly, with scientific writing being the hardest to detect. The experiment tested three tools: Pangram, GPTZero, and Originality.ai. It used 495 human-written texts covering blogs, novels, and scientific writing (all created before ChatGPT existed), and found that style imitation can effectively evade detection.
SenseTime open-sources SenseNova-Vision-7B-MoT, a multi-task vision model integrating object detection, OCR, depth estimation, normal estimation, image segmentation, and multi-view processing in a 7B architecture, providing an efficient base for visual understanding and GUI agents.....
Tian Yonglong, a former researcher at OpenAI, has joined Tencent's Large Language Model Department, focusing on the development of visual language models. This move is seen as a key recruitment for Tencent to strengthen its multi-modal large model strategy, highlighting the intense competition for cutting-edge talent.
Anthropic launches its new mid-to-high-end model, Claude Sonnet 5, focusing on cost-effectiveness. Its performance has significantly approached the flagship Opus series. The model features the strongest agent capabilities to date, enabling it to independently plan complex tasks, self-check outputs, and flexibly utilize external tools such as browsers and terminals. It performs exceptionally well in reasoning, programming, and knowledge tasks.