Skip to content

Model Training

Topic archive7 matches

Back to homeGEO summary endpoint

2026-09-08

Technology

  • Alibaba releases Qwen-Drive 1.0 AI model for autonomous driving: Alibaba's research arm has released Qwen-Drive 1.0, an AI model that integrates environmental perception, traffic Q&A, and route planning. Researchers noted that text-image models require intentional training for spatial awareness, as they do not automatically comprehend 3D space.

    Autonomous VehiclesThe Decoder

    Permalink

2026-09-07

Technology

  • Researchers train Iris search agents using web hyperlink structures: Researchers have introduced Iris-mini and Iris-pro, two search agents trained at the 35B-A3B and 397B-A17B scales. The models were trained using a data pipeline that reverse-constructs multi-hop question chains from the hyperlink structure of a web corpus.

    AI ModelsarXiv

    Permalink
  • New inference-time method reduces hallucinations in OpenAI's Whisper: Researchers have developed a training-free, inference-time method to reduce hallucinated transcripts in OpenAI's Whisper model. The approach estimates a compact hallucination-associated subspace from non-speech calibration data and projects decoder hidden states away from it.

    AI ResearcharXiv

    Permalink
  • More capable LLMs can increase systemic risk in financial markets: A study shows that improving individual large language model capability can degrade system-level outcomes rather than improve them. Researchers hypothesize that shared training and architectures cause more capable LLMs to behave similarly, creating correlated actions that do not diversify away.

    ResearcharXiv

    Permalink