Comments (1)
Testing Done:
...
File "/juice/scr/wuzhengx/align-transformers/models/mlp/modelings_mlp.py", line 60, in forward
self.act(
File "/u/nlp/anaconda/main/anaconda3/envs/wuzhengx-bootleg/lib/python3.8/site-packages/torch/nn/modules/module.py", line 1545, in _call_impl
hook_result = hook(self, args, kwargs, result)
File "/juice/scr/wuzhengx/align-transformers/models/alignable_base.py", line 604, in hook_callback
selected_output = self._gather_intervention_output(
File "/juice/scr/wuzhengx/align-transformers/models/alignable_base.py", line 392, in _gather_intervention_output
selected_output = gather_neurons(
File "/juice/scr/wuzhengx/align-transformers/models/modeling_utils.py", line 284, in gather_neurons
tensor_output = torch.gather(
(Triggered internally at ../torch/csrc/autograd/python_anomaly_mode.cpp:114.)
Variable._execution_engine.run_backward( # Calls into the C++ engine to run the backward pass
....
----------------------------------------------------------------------
Ran 13 tests in 4.751s
OK
from pyvene.
Related Issues (20)
- [P1] Adding tests for functions in `modeling_utils.py`
- [P1] Adding tests for interventions and their util functions
- [P2] Mac chip MPS mode support HOT 1
- [P1] Tutorial of Inference-time Intervention HOT 1
- Bug in BoundlessRotatedSpaceIntervention HOT 1
- [P1] Speed up training of multiple DAS interventions with caching
- [P2] Sparse autoencoders HOT 3
- [P1] Optionally remove the dependency of the config file HOT 1
- [P0] Upgrade tutorials to new API HOT 1
- [P1] Dynamic Intervention Scheduler
- [P2] Add a new huggingface collator for working with Pyvene models
- [P0] Make interface compatible with HF trainer
- [Bug]: pyvene.ai is taking too long to respond HOT 2
- [P0] Adding back GPT2 and other model supports
- [Bug]: Is this repo support MPT architecture ? i got error HOT 1
- [P0] `IntervenableModel.save()` doesn't save trained model parameters
- [P1] Tuned lens support
- [Bug]: Whether to support intervene_on_prompt=False HOT 2
- [Feature Request / Suggestion]: Capturing of the residual stream HOT 5
- [P0] Making `RepresentationConfig` a PretrainedConfig object
Recommend Projects
-
React
A declarative, efficient, and flexible JavaScript library for building user interfaces.
-
Vue.js
🖖 Vue.js is a progressive, incrementally-adoptable JavaScript framework for building UI on the web.
-
Typescript
TypeScript is a superset of JavaScript that compiles to clean JavaScript output.
-
TensorFlow
An Open Source Machine Learning Framework for Everyone
-
Django
The Web framework for perfectionists with deadlines.
-
Laravel
A PHP framework for web artisans
-
D3
Bring data to life with SVG, Canvas and HTML. 📊📈🎉
-
Recommend Topics
-
javascript
JavaScript (JS) is a lightweight interpreted programming language with first-class functions.
-
web
Some thing interesting about web. New door for the world.
-
server
A server is a program made to process requests and deliver data to clients.
-
Machine learning
Machine learning is a way of modeling and interpreting data that allows a piece of software to respond intelligently.
-
Visualization
Some thing interesting about visualization, use data art
-
Game
Some thing interesting about game, make everyone happy.
Recommend Org
-
Facebook
We are working to build community through open source technology. NB: members must have two-factor auth.
-
Microsoft
Open source projects and samples from Microsoft.
-
Google
Google ❤️ Open Source for everyone.
-
Alibaba
Alibaba Open Source for everyone
-
D3
Data-Driven Documents codes.
-
Tencent
China tencent open source team.
from pyvene.