
Continued pre-training on unlabeled OpenWebSearch.eu data improved BERT models by ~3% average macro-F1 across 16 benchmarks, with larger gains in low-resource languages
Ensemble of four open-source LLMs (Mistral-7B, Llama3.1-8B, Gemma2-9B, Qwen2.5-14B) generated synthetic annotations for hate speech detection
LightGBM meta-learner ensemble outperformed simpler strategies like mean averaging and majority voting for combining LLM predictions
Study covers English, German, Spanish, and Vietnamese languages, demonstrating improved cross-lingual generalization for hateful content detection
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Goldman Sachs, Morgan Stanley and Citigroup are pressing elite law firms to lower fees, arguing AI is sharply…

MediaTek shares closed 10% higher on Tuesday after the Taiwanese chip firm announced a partnership with Nvidia

Walmart settled opioid dispensing claims for $50 million

GE Vernova (NYSE: GEV) announced on August 24, 2026, in Paris, a new medium-voltage uninterruptible power supp…

Dell raised its annual revenue outlook, citing strong sales of AI servers

Oracle (ORCL) stock is down 4% to $142.82 in Tuesday afternoon trading, after the 10-year Treasury yield climb…
