top of page


PROCESSBENCH: Toward a Scalable Evaluation of Mathematical Reasoning Errors in AI
PROCESSBENCH evaluates step-by-step errors in AI models, enhancing analysis and oversight in solving complex math problems.

Andrea Viliotti
16 dic 2024Tempo di lettura: 10 min


AI Knowledge Circuits
AI Knowledge Circuits: Transformers use internal circuits to manage and update knowledge with targeted, efficient interventions.

Andrea Viliotti
15 dic 2024Tempo di lettura: 7 min


EvalGIM: a unified platform for evaluating generative image models
EvalGIM evaluates text-to-image models on quality, diversity, and consistency. Modular and flexible, it supports research and business.

Andrea Viliotti
15 dic 2024Tempo di lettura: 8 min


Technology 2025: Evolving Global Dynamics
Technology 2025: AI and AR transform markets and society. Innovation and regulation must balance to mitigate global risks.

Andrea Viliotti
15 dic 2024Tempo di lettura: 8 min


Tech Trends 2025. Artificial Intelligence, the Cognitive Substrate for the Digital Future
Tech Trends 2025: AI is the cognitive substrate of the digital future, requiring ethics, transparency, and integration.

Andrea Viliotti
13 dic 2024Tempo di lettura: 9 min


RARE: optimizing LLM reasoning
RARE boosts LLM reasoning with A6, A7 retrieval and RAFS. Open-source models exceed GPT-4. Limits: cost and non-standard metrics.

Andrea Viliotti
10 dic 2024Tempo di lettura: 6 min


Introduction to Machine Learning
Machine Learning improves tasks by learning from data, requiring ethical management to reduce biases and social impacts.

Andrea Viliotti
5 dic 2024Tempo di lettura: 15 min


AI in Italy: From Benefits for Large Companies to Challenges for SMEs
AI in Italy: €760M market, 90% large companies’ investments, SMEs 18%. Barriers: skills, costs. Training and incentives needed.

Andrea Viliotti
4 dic 2024Tempo di lettura: 10 min


Challenges and Opportunities in Deploying Generative AI Systems
Generative AI is growing, but integrating solutions remains difficult. Trust and speed are keys to success.

Andrea Viliotti
3 dic 2024Tempo di lettura: 9 min


Boston Consulting Group. The AI Maturity Map
Boston Consulting Group: AI maturity
The USA and Singapore lead thanks to investments and ethics.

Andrea Viliotti
3 dic 2024Tempo di lettura: 14 min


Corporate AI. Analysis of the Cisco AI Readiness Index 2024
Corporate AI. Only 10% of Italian companies are leaders. Challenges include inadequate infrastructure, fragmented data, and talent gaps.

Andrea Viliotti
3 dic 2024Tempo di lettura: 8 min


Impact of AI in Accounting and Finance
AI in accounting and finance optimizes processes and decisions but needs governance and quality data for sustainable strategic results.

Andrea Viliotti
2 dic 2024Tempo di lettura: 8 min


Multimodal AI in Medicine
Multimodal AI combines data for personalized care. Ethical and privacy challenges require innovation and strategic vision for success.

Andrea Viliotti
1 dic 2024Tempo di lettura: 10 min


OWASP Top 10 LLM: Ten Vulnerabilities for LLM-Based Applications
OWASP Top 10 LLM 2025 lists key vulnerabilities and measures for LLMs: input filters, human supervision, and monitoring for safety.

Andrea Viliotti
1 dic 2024Tempo di lettura: 20 min


Gaming and Artificial Intelligence. BALROG the New Standard for LLMs and VLMs
BALROG evaluates LLMs/VLMs in complex environments, testing reasoning and planning. It seeks to address agentic and multimodal limitations.

Andrea Viliotti
30 nov 2024Tempo di lettura: 9 min


GenAI in Banking
GenAI in banking transforms services and risk management through gradual adoption. Governance, ethics, and security are crucial for success.

Andrea Viliotti
28 nov 2024Tempo di lettura: 9 min


LLMs and Security: MRJ-Agent for a Multi-Round Attack
MRJ-Agent uses risk decomposition and psychological induction for effective multi-round attacks against advanced AI model defenses.

Andrea Viliotti
28 nov 2024Tempo di lettura: 10 min


BrainBench: Language Models Surpass Neuroscience Experts
BrainBench reveals LLMs predict scientific results with 81.4% accuracy, surpassing human experts at 63.4%.

Andrea Viliotti
28 nov 2024Tempo di lettura: 12 min


Configurable Foundational Models: A Modular Approach to Building LLMs
Configurable Foundational Models: dynamic LLM modules, more scalable, adaptable, and ideal for efficiency and personalization.

Andrea Viliotti
17 nov 2024Tempo di lettura: 14 min
bottom of page