Fetching the paper…
Reading the bibliography…
Frontier AI safety policies highlight automation of AI research and development (R&D) by AI agents as an important capability to anticipate.
Chip placement with deep reinforcement learning (2020)
Mirhoseini, A. et al · 2004
Earlier work this paper cites.
Evaluating large language models trained on code
Chen, M. et al · 2021
Earlier work this paper cites.
Competition-level code generation with AlphaCode
Li, Y. et al · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback (2022)
Ouyang, L. et al · 2022
Earlier work this paper cites.
Constitutional AI: Harmlessness from AI feedback (2022)
Bai, Y. et al · 2022
Earlier work this paper cites.
Textbooks are all you need (2023)
Gunasekar, S. et al · 2023
Earlier work this paper cites.
OpenAI preparedness framework (beta)
OpenAI · 2023
Earlier work this paper cites.
SWE-bench: Can language models resolve real-world github issues?
Jimenez, C. E. et al · 2023
Earlier work this paper cites.
GPQA: A graduate-level google-proof q&a benchmark
Rein, D. et al · 2023
Earlier work this paper cites.
GAIA: a benchmark for general AI assistants (2023)
Mialon, G. et al · 2023
Earlier work this paper cites.
Evoprompting: Language models for code-level neural architecture search (2023)
Chen, A., Dohan, D. M. & So, D. R · 2023
Earlier work this paper cites.
Chemcrow: Augmenting large-language models with chemistry tools (2023)
Bran, A. M. et al · 2023
Earlier work this paper cites.
mosaicml/llm-foundry (2023)
King, D. et al · 2023
Earlier work this paper cites.
The Llama 3 herd of models (2024)
AI at Meta · 2024
Earlier work this paper cites.
Self-alignment with instruction backtranslation
Li, X. et al · 2024
Earlier work this paper cites.
OpenDevin: An open platform for AI software developers as generalist agents (2024)
Wang, X. et al · 2024
Earlier work this paper cites.
Interviewing AI researchers on automation of AI r&d (2024)
Owen, D · 2024
Earlier work this paper cites.
Estimating idea production: A methodological survey
Erdil, E., Besiroglu, T. & Ho, A · 2024
Earlier work this paper cites.
Explosive growth from AI automation: A review of the arguments (2024)
Erdil, E. & Besiroglu, T · 2024
Earlier work this paper cites.
Frontier safety framework, version 1.0
Google DeepMind · 2024
Earlier work this paper cites.
Anthropic responsible scaling policy
Anthropic · 2024
Cited alongside, same era.
Recital 110 of the eu artificial intelligence act (2024)
Union, E · 2024
Cited alongside, same era.
Memorandum on advancing the United States’ leadership in artificial intelligence; harnessing artificial intelligence to fulfill national security objectives; and fostering the safety, security, and trustworthiness of artificial intelligence (2024)
The White House · 2024
Cited alongside, same era.
The bletchley declaration on AI safety (2023)
UK Government and International Partners · 2024
Cited alongside, same era.
Frontier AI safety commitments, AI seoul summit 2024
United Kingdom Department for Science, Innovation & Technology · 2024
Cited alongside, same era.
Building an early warning system for llm-aided biological threat creation (2024)
MLR-copilot: Autonomous machine learning research based on large language models agents (2024)
Li, R., Patel, T., Wang, Q. & Du, X · 2024
Closest in time.
Guo, S. et al · 2024
Closest in time.
SUPER: Evaluating agents on setting up and executing tasks from research repositories (2024)
Bogin, B. et al · 2024
Closest in time.
The AI scientist: Towards fully automated open-ended scientific discovery (2024)
Lu, C. et al · 2024
Closest in time.
Eureka: Human-level reward design via coding large language models (2024)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
OpenAI · 2024
Cited alongside, same era.
Wan, S. et al · 2024
Cited alongside, same era.
What a compute-centric framework says about takeoff speeds (2023)
Davidson, T · 2024
Cited alongside, same era.
Sabotage evaluations for frontier models (2024)
Benton, J. et al · 2024
Cited alongside, same era.
Securing AI model weights: Preventing theft and misuse of frontier models
Nevo, S. et al · 2024
Cited alongside, same era.
Raising the bar on swe-bench verified with claude 3.5 sonnet (2024)
Anthropic · 2024
Cited alongside, same era.
Artificial intelligence index report 2024 (2024)
Maslej, N. et al · 2024
Cited alongside, same era.
Ma, Y. J. et al · 2024
Closest in time.
Faldor, M., Zhang, J., Cully, A. & Clune, J · 2024
Closest in time.
Discovering preference optimization algorithms with and for large language models (2024)
Lu, C. et al · 2024
Closest in time.
Sciagent: Tool-augmented language models for scientific reasoning (2024)
Ma, Y. et al · 2024
Closest in time.
Scientific large language models: A survey on biological & chemical domains (2024)
Zhang, Q. et al · 2024
Closest in time.
Tornede, A. et al · 2024
Closest in time.
Trends in machine learning hardware (2023)
Hobbhahn, M., Heim, L. & Aydos, G · 2024
Closest in time.
Vivaria: Open-source platform for agent evaluations (2024)
METR · 2024
Closest in time.
Claude 3.5 Sonnet model card addendum (2024)
Anthropic · 2024
Closest in time.
OpenAI o1 system card (2024)
OpenAI · 2024
Closest in time.
AIDE: Data science automation technical report (2024)
weco.ai · 2024
Closest in time.
Details about METR’s preliminary evaluation of OpenAI o1-preview (2024)
METR · 2024
Closest in time.
Claude 3.5 Sonnet: Quality, performance & price analysis (2024)
Analysis, A · 2024
Closest in time.
Training compute-optimal large language models
Hoffmann, J. et al · 2024
Closest in time.
Direct preference optimization: Your language model is secretly a reward model (2024)
Rafailov, R. et al · 2024
Closest in time.