• About Us
  • Disclaimer
  • Contact Us
  • Privacy Policy
Thursday, July 3, 2025
mGrowTech
No Result
View All Result
  • Technology And Software
    • Account Based Marketing
    • Channel Marketing
    • Marketing Automation
      • Al, Analytics and Automation
      • Ad Management
  • Digital Marketing
    • Social Media Management
    • Google Marketing
  • Direct Marketing
    • Brand Management
    • Marketing Attribution and Consulting
  • Mobile Marketing
  • Event Management
  • PR Solutions
  • Technology And Software
    • Account Based Marketing
    • Channel Marketing
    • Marketing Automation
      • Al, Analytics and Automation
      • Ad Management
  • Digital Marketing
    • Social Media Management
    • Google Marketing
  • Direct Marketing
    • Brand Management
    • Marketing Attribution and Consulting
  • Mobile Marketing
  • Event Management
  • PR Solutions
No Result
View All Result
mGrowTech
No Result
View All Result
Home Al, Analytics and Automation

University of Michigan Researchers Propose G-ACT: A Scalable Machine Learning Framework to Steer Programming Language Bias in LLMs

Josh by Josh
June 30, 2025
in Al, Analytics and Automation
0
University of Michigan Researchers Propose G-ACT: A Scalable Machine Learning Framework to Steer Programming Language Bias in LLMs
0
SHARES
0
VIEWS
Share on FacebookShare on Twitter


LLMs and the Need for Scientific Code Control

LLMs have rapidly evolved into complex natural language processors, enabling the development of agentic systems that manage complex workflows. However, the use of LLM agents for generating scientific code is unexplored. Scientific software primarily depends on C++, CUDA, and other low-level languages, which are underrepresented in most pretraining datasets. As a result, implementations generated by LLMs contain syntactic or semantic errors, which lead to compilation issues or unstable runtime behavior. Existing agents rely heavily on user-specified control primitives and carefully crafted prompts, which are prone to misinterpretation and can lead to erratic execution flows.

Limitations of Existing Steering Methods

Recent approaches have been developed to tackle LLM steering challenges by uncovering causal links within model activations and facilitating precise neuron-level interventions. SFT, weight modulation techniques, and RLHF represent direct intervention for model steering, but they have significant computational overhead and may reduce the model’s robustness and general performance. Activation Patching, which uses corrupted inputs as a baseline distribution, is widely adopted for fine-grained output control. However, these methods demand extensive model sweeps involving millions of evaluations and are used on multiple-choice question benchmarks, rather than real-world deployment scenarios.

Introduction of G-ACT Framework

Researchers from the University of Michigan have proposed a gradient-refined adaptive activation steering framework (G-ACT) to address the challenge of steering scientific code generation toward specific programming languages in LLMs. It arises from evaluating five causal LLMs on scientific coding prompts. G-ACT clusters per-prompt activation differences into steering directions and uses lightweight per-layer probes that are trained and refined online to select suitable steering vectors. The framework supports concept-level control while ensuring scalability and interpretability, providing a practical method for achieving reproducible behavior in agentic systems that require consistent programming language choices for scientific computing tasks.

Model Evaluation and Baseline Biases

Researchers evaluate five instruction-tuned LLMs, including Llama-3.2-3B-Instruct, Llama-3.3-70B-Instruct, Qwen2.5-Coder-32B-Instruct, Qwen2.5-14B-Instruct-1M, and QwQ-32B. Each model is tested on 84 benchmark questions with 25 repetitions per prompt at sampling temperature 1.0 to ensure statistical stability. Results for language preferences reveal that Llama-3.2-3B strongly defaults to Java (76.2%), while Llama-3.3-70B favors Python (73.8%). Qwen models show different biases with Qwen2.5-Coder preferring Python (59.5%) and Qwen2.5-14B favoring Julia (66.7%). These baseline measurements show that model scale, architectural design, and fine-tuning data collectively create reproducible biases.

Static Neuron Activation and Language Biasing

Static method analysis involves inducing language preference bias and code generation testing. Results for preference bias show that selective activation of individual MLP neurons in baseline tests with Llama-3.2-3B-Instruct gains strong causal control over programming language selection. When targeting CPP generation, results show nearly 100% CPP output across most problems, virtually eliminating Python, Java, and Julia outputs. Moreover, code generation testing reveals two distinct behavioral regimes: Python-leaning tasks show 40-80% Python outputs for high-level operations, while CPP-dominant tasks exhibit 60-90% CPP preference for performance-critical routines. The model achieves ~73% CPP generation more often than Python, but still defaults to Python for a significant portion of prompts.

Gradient-Refined Activation Steering Results

In this paper, researchers present a gradient-refined adaptive activation steering that can control programming language selection in scientific code generation. The framework achieves substantial improvements, increasing probe classification accuracy from 0% to 61.5% in early layers of LLaMA-3.2 3B. Despite a modest runtime overhead of 1.3-1.4 times slower generation, the framework remains practical through selective layer steering and caching optimizations. G-ACT offers a scalable and interpretable approach for concept-level control that goes beyond programming languages by embedding persistent transformation matrices. This ensures consistent model behavior across users and introduces a new standard for reliable LLM steering in scientific computing contexts.


Check out the Paper. All credit for this research goes to the researchers of this project. Also, feel free to follow us on Twitter and don’t forget to join our 100k+ ML SubReddit and Subscribe to our Newsletter.


Sajjad Ansari is a final year undergraduate from IIT Kharagpur. As a Tech enthusiast, he delves into the practical applications of AI with a focus on understanding the impact of AI technologies and their real-world implications. He aims to articulate complex AI concepts in a clear and accessible manner.



Source_link

READ ALSO

Confronting the AI/energy conundrum

Baidu Open Sources ERNIE 4.5: LLM Series Scaling from 0.3B to 424B Parameters

Related Posts

Confronting the AI/energy conundrum
Al, Analytics and Automation

Confronting the AI/energy conundrum

July 3, 2025
Baidu Open Sources ERNIE 4.5: LLM Series Scaling from 0.3B to 424B Parameters
Al, Analytics and Automation

Baidu Open Sources ERNIE 4.5: LLM Series Scaling from 0.3B to 424B Parameters

July 2, 2025
Novel method detects microbial contamination in cell cultures | MIT News
Al, Analytics and Automation

Novel method detects microbial contamination in cell cultures | MIT News

July 2, 2025
Baidu Researchers Propose AI Search Paradigm: A Multi-Agent Framework for Smarter Information Retrieval
Al, Analytics and Automation

Baidu Researchers Propose AI Search Paradigm: A Multi-Agent Framework for Smarter Information Retrieval

July 2, 2025
Merging design and computer science in creative ways | MIT News
Al, Analytics and Automation

Merging design and computer science in creative ways | MIT News

July 1, 2025
Building Advanced Multi-Agent AI Workflows by Leveraging AutoGen and Semantic Kernel
Al, Analytics and Automation

Building Advanced Multi-Agent AI Workflows by Leveraging AutoGen and Semantic Kernel

July 1, 2025
Next Post
Does ChatGPT suffer? If AI becomes conscious, it could.

Does ChatGPT suffer? If AI becomes conscious, it could.

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

POPULAR NEWS

Communication Effectiveness Skills For Business Leaders

Communication Effectiveness Skills For Business Leaders

June 10, 2025
7 Best EOR Platforms for Software Companies in 2025

7 Best EOR Platforms for Software Companies in 2025

June 21, 2025
Eating Bugs – MetaDevo

Eating Bugs – MetaDevo

May 29, 2025
Top B2B & Marketing Podcasts to Lead You to Succeed in 2025 – TopRank® Marketing

Top B2B & Marketing Podcasts to Lead You to Succeed in 2025 – TopRank® Marketing

May 30, 2025
Entries For The Elektra Awards 2025 Are Now Open!

Entries For The Elektra Awards 2025 Are Now Open!

May 30, 2025

EDITOR'S PICK

HP reveals $24,999 hardware created just for Google Beam

HP reveals $24,999 hardware created just for Google Beam

June 13, 2025
Why 3 Podcast Interviews Might Be More Valuable Than 300 LinkedIn Posts

Why 3 Podcast Interviews Might Be More Valuable Than 300 LinkedIn Posts

June 20, 2025
9 conseils d’expert pour vos scénarios marketing automation

9 conseils d’expert pour vos scénarios marketing automation

May 30, 2025

Best Shopify Marketing Apps for 2025 –

June 3, 2025

About

We bring you the best Premium WordPress Themes that perfect for news, magazine, personal blog, etc. Check our landing page for details.

Follow us

Categories

  • Account Based Marketing
  • Ad Management
  • Al, Analytics and Automation
  • Brand Management
  • Channel Marketing
  • Digital Marketing
  • Direct Marketing
  • Event Management
  • Google Marketing
  • Marketing Attribution and Consulting
  • Marketing Automation
  • Mobile Marketing
  • PR Solutions
  • Social Media Management
  • Technology And Software
  • Uncategorized

Recent Posts

  • The 8 Best AI Detectors, Tested and Compared
  • The 7 best smartwatches for Android in 2025
  • Floki Launches Norse-Themed Blockchain Game with Real Rewards
  • Even before the Xbox layoffs, there was ‘tension’ at Halo Studios
  • About Us
  • Disclaimer
  • Contact Us
  • Privacy Policy
No Result
View All Result
  • Technology And Software
    • Account Based Marketing
    • Channel Marketing
    • Marketing Automation
      • Al, Analytics and Automation
      • Ad Management
  • Digital Marketing
    • Social Media Management
    • Google Marketing
  • Direct Marketing
    • Brand Management
    • Marketing Attribution and Consulting
  • Mobile Marketing
  • Event Management
  • PR Solutions

Are you sure want to unlock this post?
Unlock left : 0
Are you sure want to cancel subscription?