← Back to all news
OpenAI Blog July 25, 2022

A hazard analysis framework for code synthesis large language models

Read the original on OpenAI Blog →

The Flow has not summarised this story yet — read it at OpenAI Blog.

  • llms

Related stories

arXiv Machine Learning
Aug 3

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

arXiv:2607. 29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications.

By Jan Marius St\"urmer, Jascha Knack, Tobias Koch, Andreas Weinmann
llmsbenchmarks
More like this →
OpenAI Blog
Jul 7, 2021

Evaluating large language models trained on code

llms
More like this →
arXiv AI
Jul 3

TokenScope: Token-Level Explainability and Interpretability for Code-Oriented Tasks in Large Language Models

arXiv:2607. 01235v1 Announce Type: cross Abstract: Understanding how Large Language Models (LLMs) make token-level decisions during code generation remains a major challenge for both researchers and practitioners.

By Amirreza Esmaeili, Fatemeh Fard
llmssafety
More like this →
arXiv AI
Jul 14

Minionese: Comprehensive Benchmark and Mechanistic Study of Multilingual LLM Safety

arXiv:2607. 10112v1 Announce Type: cross Abstract: Safety alignment in large language models remains brittle across languages: prompts reliably refused in English can elicit harmful compliance in non-English and low-resource settings.

By Chigozirim Ifebi, Brent Kong, Ayushi Mehrotra
llmsbenchmarkssafety
More like this →
arXiv AI
2d ago

The Fools are Certain; the Wise are Doubtful: Exploring LLM Confidence in Code Completion

arXiv:2508. 16131v3 Announce Type: replace-cross Abstract: Code completion entails the task of providing missing tokens given a surrounding context.

By Zoe Kotti, Konstantina Dritsa, Diomidis Spinellis, Panos Louridas
llmsfine-tuningsafety
More like this →
OpenAI Blog
Mar 3, 2022

Lessons learned on language model safety and misuse

We describe our latest thinking in the hope of helping other AI developers address safety and misuse of deployed models.

llmssafety
More like this →