📁 LLM

The LLM archive monitors model releases, quantization updates, reasoning capabilities, and real-world deployment implications for local and hybrid AI. We focus on what materially changes selection and operations: context windows, latency, memory footprint, licensing, and evaluation evidence across open and commercial families. This section is designed for teams that need dependable model intelligence, not hype cycles. Pair these updates with the LLM pillar and references to hardware constraints and framework integration.

A user shares their experience with local language models, highlighting the accelerated learning curve compared to using cloud solutions. The article touches on topics such as context optimization, KV cache management, and exploration of Mixture of Experts architectures.

2026-02-27 Fonte

A new study explores the use of LLMs, specifically GPT-5, for analyzing the context of textual citations. The research focuses on prompt sensitivity, varying their structure to assess how they influence the model's interpretations. The goal is to understand if LLMs can support complex interpretative analyses.

2026-02-27 Fonte

A novel framework, Decoder-based Sense Knowledge Distillation (DSKD), integrates structured lexical resources into the training of decoder-style large language models (LLMs). This approach enhances performance without requiring dictionary lookups at inference time, enabling generative models to inherit structured semantics while maintaining efficient training.

2026-02-27 Fonte

A novel passive surveillance system, powered by artificial intelligence and graph neural networks, aims to detect early stroke risk in high-risk individuals by analyzing patient-reported symptoms. The approach combines a symptom taxonomy with a machine learning model to identify predictive patterns.

2026-02-27 Fonte

FIRE is a new benchmark for evaluating LLM capabilities in the financial domain. It includes theoretical knowledge tests based on certification exams and practical scenarios with 3,000 questions. Results obtained with state-of-the-art models, such as XuanYuan 4.0, have been made public to foster research.

2026-02-27 Fonte

A new system, GYWI, combines author knowledge graphs with retrieval-augmented generation (RAG) to provide controllable academic context and traceable inspiration pathways for large language models (LLMs) in generating new scientific ideas. The system was evaluated with different LLMs, including GPT-4o and DeepSeek-V3.

2026-02-27 Fonte

Google has unveiled Nano Banana 2, an artificial intelligence model for image editing. The model appears capable of altering the reality of photos, opening up new creative possibilities, albeit with sometimes unpredictable results. An analysis of the capabilities and limitations of this tool.

2026-02-27 Fonte

According to the ORCA test, current large language models (LLMs), while improving, remain prediction engines and do not always provide the correct solution to mathematical problems. Even Gemini 3 Flash, among the top performers, would receive a mediocre grade.

2026-02-26 Fonte

A recent Fortune study reveals that AI-powered search engines are confidently wrong over 60% of the time. This is also reflected in automatically generated captions, often filled with errors and incomprehensible. The article explores the implications of this unreliability for the accessibility and usability of information.

2026-02-26 Fonte

Anthropic has launched a new marketing blog, attributing the writing to Claude Opus 3, a "retired" large language model (LLM). The initiative raises questions about the role and perception of LLMs in marketing and communication.

2026-02-26 Fonte

Read AI introduces Ada, a digital assistant integrated into emails. Ada manages your availability and provides answers based on the company's knowledge base and information from the web, simplifying communication and scheduling.

2026-02-26 Fonte

Read AI is launching Ada, an email-integrated digital assistant. Ada is designed to automate schedule management and provide quick answers, drawing from both internal company knowledge and external web resources, thereby enhancing user productivity.

2026-02-26 Fonte

OpenAI and the Pacific Northwest National Laboratory (PNNL) introduce DraftNEPABench, a benchmark for evaluating how AI agents can accelerate federal permitting, potentially reducing drafting time by up to 15% and modernizing infrastructure reviews.

2026-02-26 Fonte

Google has announced Nano Banana 2, a new version of its AI model focused on image generation. The model will be integrated as the default option in the Gemini app and in AI mode, promising superior performance compared to the previous version.

2026-02-26 Fonte

A recent BBC article explored how generative AI tools could be "hacked" within minutes by introducing newly published online content. The original article suggests that AI models like ChatGPT can be easily influenced by unverified information, raising questions about their reliability. However, this view may not consider the complexity and countermeasures in place.

2026-02-26 Fonte