Research · 研究

Pushing the frontier of reasoning, code, and Chinese-language AI.

We work on AI safety, agentic coding, multimodal evaluation, reasoning training, low-resource NLP, and model auditing, writing each result up as a bilingual explainer for researchers and general readers alike.


Reasoning · Code · Chinese-language AIUpdated June 2026
Featured work · AI safety

SafeGEO measures how recommendation agents can be steered — and defended.

A controlled benchmark of search-optimization attacks and mitigations across real recommendation settings.

Read the research
600Recommendation cases 22Attack variants EMNLP2026
Milestones

From one spotlight to five research lines in a year.

Three stages: a first paper selected as a NeurIPS SoLaR Spotlight in 2024, the lab founded and SEAM accepted at COLM in 2025, and five research lines released across 2026 — two of them at a CCF-A main conference and EMNLP.

  1. 01Origin

    A first paper, a NeurIPS SoLaR Spotlight

    Report Cards describes model ability in natural language, then validates the descriptions themselves with contrastive accuracy and Card Elo.

    1. DecReport CardsNeurIPS SoLaR 2024
    2024 research
  2. 02Founded

    The lab is founded; SEAM accepted at COLM 2025

    The same question in text or in image form often draws different answers. SEAM quantifies cross-modal consistency across 21 models, four domains and 9,600 evaluations.

    1. JunCoolwei AI Lab founded
    2. AugSEAMCOLM 2025
    3. AugMobile-Agent-BenchProject launched
    2025 research
  3. 03Ongoing

    Five research lines released

    In a single year the work extends from agent benchmarking to reasoning training, master distillation and multilingual datasets, with two results at a CCF-A main conference and EMNLP.

    1. FebSWE-Bench MobileKDD 2026
    2. MarOasisSimpLow-resource simplification dataset
    3. MarGrounded Chess ReasoningMaster distillation
    4. AprThinkTwiceSelf-refinement RLVR
    5. OctSafeGEOEMNLP 2026
    2026 research