• New accounts will normally be approved within 24 hours.
    Note: We're currently having issues with our e-mail system, anything requiring e-mail validation (2FA, forgotten passwords, etc.) must be changed manually by an Admin. Please reach out via the Contact Us form if you require any assistance.

Other Software Maghni AI by Crescendia

sunnyp4rk

stuck in the computer
Jan 23, 2020
523
20
Midwest US
after-rain.net
I would love to know what went through VocaTone's mind when they thought that making new versions of their voicebanks for an engine that would have no release date for years would be a good idea. Maybe they thought it would be done by now, but they should really just move a lot of it to Synth V or V6 or any other functional engine. I'd love to see Oliver be revived, but not in a state like this.
 

Blue Of Mind

The world that I do not know...
Apr 8, 2018
884
I've gotten the impression they may not have had another choice - Vocatone was always kinda messy and I suspect they burned bridges that way with a lot of third parties in the vocal synth scene (particularly Yamaha)
As an old fart, I can chime in for any younger fans on here that VocaTone has long had a reputation for being dickish, up to and including what @DefiantKitsune has said - and I can back this up, because I remember VT posting on VocaloidOtaku when I was a member there, and they eventually managed to alienate a good chunk, if not the entirety of the VO community with their messiness.

If you've ever wondered why their banks haven't simply been ported to V6 or even SynthV, there's probably a very good reason why. 👀
 
Feb 18, 2025
18
So all hell has broken loose as the new head of Vocatone, Crossy, already has accusations of being a shotacon

Meanwhile, Cocoa has decided to share what some actual magni renders sound like here

Not sure if this really counters or explains anything further, but it's interesting to have an idea of what they were working with.

Cocoa also seems determined that they're still all making a commercial product

1000025601.jpg

Edit: oops missed a page, sorry for reposting the demos
 
Last edited:

SeleDreams

Hardcore Fan
Jul 31, 2019
336
25
So all hell has broken loose as the new head of Vocatone, Crossy, already has accusations of being a shotacon

Meanwhile, Cocoa has decided to share what some actual magni renders sound like here

Not sure if this really counters or explains anything further, but it's interesting to have an idea of what they were working with.
I find it funny to see this as a big deal when that's like... a big chunk of the most popular japanese vocaloid producers.
 
  • Like
Reactions: StrawbebbyPancakes

lIlI

Staff member
Administrator
Apr 6, 2018
1,273
Kanru's walls
As it's understandably one of the most contentious parts of this controversy, I've compiled all the statements regarding the use of RVC, Melodyne, and Synthesizer V in the creation of a Maghni AI demo for posterity.

The original screenshot of the Maghni devs discussing enhancing the demo behind the scenes, the authenticity of which has not been disputed by any party:



In this screenshot Suzie (demo artist) suggests that they splice in phonemes to improve the demo. Cocoa (Crescendia developer) agrees, while acknowledging that it is unethical. Editing individual phonemes in Melodyne is also discussed. Suzie later clarifies that she used Synthesizer V output, then converted it using RVC to create the spliced samples.

In her Twitter thread about the screenshot, Suzie says that consonants and a spoken voice line were generated with RVC (RVC is an open source AI voice changer, not associated with Maghni).


Austin Grissom's statement corroborates the use of samples from other engines, and adds that a Synthesizer V SVP file was used for the timing. He states that audio output from Synthesizer V was not one of the places they spliced from.



Natalia confirms that the demos were heavily processed, and corroborates that RVC was used. She confirms that RVC and 'related processing' were used for selection of consonants, and the spoken word portion.


And finally, Cocoa confirms again that consonants were spliced and RVC was used to create the spoken word section, and that Synthesizer V used as an editor. However, he disputes Synthesizer V being used for phonetic timing.



Suzie contradicts Cocoa and Austin to say that audio output from Synthesizer V was used by her to produce audio samples which were then fed into RVC.

In their Q&A, the current Maghni team (Cinna, Aku-P, Cocoa, and Tart) state that RVC was used splice high notes, consonants, and the speaking section. They describe the demo as 60-80% Maghni.


Their description of the use of Synthesizer V denies its use for generating audio:


All these stories, with the exception of exactly how SynthV was used, reinforce each other, and are consistent with Suzie's first statement. It's extremely clear what happened: I assume any confusion on social media comes from people skim-reading statements, or jumping to conclusions based on vaguetweets and secondhand summaries.

What's worrying is that Suzie's demo doesn't seem to have been the only misleading demonstration.

Cocoa and Suzie's final conversation implies that consonants from RVC or real singing audio were spliced into multiple demos. (Suzie was in charge of splicing on Eggtan's demo, whereas Cocoa recalls his own separate splicing tasks) Edit: The use of splicing in multiple demos is currently denied by Crescendia, saying that two potential spliced versions were made of the same demo by Cocoa and Suzie separately.

I stand by my initial opinion that this usage was unethical (especially if audio output from Synthesizer V was used in the process, as Suzie alleges), a sentiment I know the Maghni AI devs publicly and privately concur with (ahah). Now the spotlight is on them, I hope they'll be extra rigorous about showing accurate demos. I appreciate that they're now issuing apologies, but this is one incident that's best disproved with actions.

*edit* Added the screenshot from Crescendia's Q&A!
 
Last edited:

sunnyp4rk

stuck in the computer
Jan 23, 2020
523
20
Midwest US
after-rain.net
I feel like the thing about having to use RVC and splicing to make the consonants sound right is that it's a very misleading tactic to fool people into thinking the product sounds right when it doesn't. All I'm saying is that if the software wasn't borked, you wouldn't have to pretend like it was working. They admitted to it but it sounds horrible when you think about it. Even using timing data is questionable because that's not representative of what the actual timing data would be like.

I still think they should cancel the project, since it's going to continue to be a giant headache for everyone involved. If fund management had been accurate and not reckless from the start, then I could excuse the length of time it's taken for Maghni to be released, since it's a small team. Now that we know what we do, it would probably be poor taste to continue development to "make a profit".

So all hell has broken loose as the new head of Vocatone, Crossy, already has accusations of being a shotacon
So I thought the art looked familiar and then I realized it's the same artist for Oliver's M.AI illustration (with some Oliver artwork in the doc). And Oliver's official age is 12. Oh boy.......

Edit:
So. The questionable Oliver stuff in that doc was worse than I thought. As an adult, I've come to realize that Oliver's existence is more questionable the more I think about it. This doc kind of solidifies that for me. Please VocaTone, let that boy rest in peace.

There was a combination of both SFW and NSFW art where Crossy called him a shota. While also working on Oliver V3 at the same time. With the child recording data.:piko_ani_lili: Awful. Originally I was going to be like "oh shotacon allegations are kind of the least most pressing matter in regards to M.AI" but....yeesh

Edit 2:
Crossy apparently made a [Admin Edit: Warning for discussion of sexual assault] bluesky post saying they made shota artwork due to abuse as a child. That's still not a great excuse to sexualize a character that had connections to a real life child with the recording data of that child. A lot of the shota art was made when Crossy was in their early 20s.

Just gonna add this bit from the doc directly in here:
Screenshot (4377).png
 
Last edited by a moderator:

SeleDreams

Hardcore Fan
Jul 31, 2019
336
25
As it's understandably one of the most contentious parts of this controversy, I've compiled all the statements regarding the use of RVC, Melodyne, and Synthesizer V in the creation of a Maghni AI demo for posterity.

The original screenshot of the Maghni devs discussing enhancing the demo behind the scenes, the authenticity of which has not been disputed by any party:



In this screenshot Suzie (demo artist) suggests that they splice in phonemes to improve the demo. Cocoa (Crescendia developer) agrees, while acknowledging that it is unethical. Editing individual phonemes in Melodyne is also discussed. Suzie later clarifies that she used Synthesizer V output, then converted it using RVC to create the spliced samples.

In her Twitter thread about the screenshot, Suzie says that consonants and a spoken voice line were generated with RVC (RVC is an open source AI voice changer, not associated with Maghni).


Austin Grissom's statement corroborates the use of samples from other engines, and adds that a Synthesizer V SVP file was used for the timing. He states that audio output from Synthesizer V was not one of the places they spliced from.



Natalia confirms that the demos were heavily processed, and corroborates that RVC was used. She confirms that RVC and 'related processing' were used for selection of consonants, and the spoken word portion.


And finally, Cocoa confirms again that consonants were spliced and RVC was used to create the spoken word section, and that Synthesizer V used as an editor. However, he disputes Synthesizer V being used for phonetic timing.



Suzie contradicts Cocoa and Austin to say that audio output from Synthesizer V was used by her to produce audio samples which were then fed into RVC.

All these stories, with the exception of exactly how SynthV was used, reinforce each other, and are consistent with Suzie's first statement. It's extremely clear what happened: I assume any confusion on social media comes from people skim-reading statements, or jumping to conclusions based on vaguetweets and secondhand summaries.

What's worrying is that Suzie's demo doesn't seem to have been the only misleading demonstration.

Cocoa and Suzie's final conversation implies that consonants from RVC or real singing audio were spliced into multiple demos. (Suzie was in charge of splicing on Eggtan's demo, whereas Cocoa recalls his own separate splicing tasks)

I stand by my initial opinion that this usage was unethical (especially if audio output from Synthesizer V was used in the process, as Suzie alleges), a sentiment I know the Maghni AI devs publicly and privately concur with (ahah). Now the spotlight is on them, I hope they'll be extra rigorous about showing accurate demos. I appreciate that they're now issuing apologies, but this is one incident that's best disproved with actions.

NGL i think that these elements would also be sufficient for a potential lawsuit from Dreamtonics, because if they did use Synthesizer V to feed into RVC (which goes against the SynthV EULA) for marketing their competing software,that's a very clear violation of dreamtonics's rights
 

lIlI

Staff member
Administrator
Apr 6, 2018
1,273
Kanru's walls
Regarding pedophilia and CSAM in relation to Crossy: due to the sensitivity of the topic and the involvement of specific individuals within the community, I think it's best we avoid any further discussion in this thread, as both sides have now said their piece.

(For clarity: It's not against forum guidelines to post abuse allegations or discuss sensitive topics with appropriate civility - I am exercising extra caution as this involves highly triggering information and risk of derailment. Thank you for understanding!)
 

lIlI

Staff member
Administrator
Apr 6, 2018
1,273
Kanru's walls
Vocatone has posted an update regarding refunds and shipments.

The craziest thing about the whole Maghni Malarkey is that it would have been avoided had Crescendia remembered to login to their Twitter account a few times, haha. By leaving it silent for a year, it resulted in backers thinking the project had been abandoned, which led to ...this. Only because I happened to see Tart post from their personal account that development was still ongoing did I know otherwise - an obscure source of news which most people naturally missed. This is perhaps the hardest learnt PR lesson I've seen from a fan project.
 

sunnyp4rk

stuck in the computer
Jan 23, 2020
523
20
Midwest US
after-rain.net
Vocatone has posted an update regarding refunds and shipments.

The craziest thing about the whole Maghni Malarkey is that it would have been avoided had Crescendia remembered to log into their Twitter account a few times, haha. By leaving it silent for a year, it resulted in backers thinking the project had been abandoned, which led to ...this. Only because I happened to see Tart post from their personal account that development was still ongoing did I know otherwise - an obscure source of news which most people naturally missed. This is perhaps the hardest learnt PR lesson I've seen from a fan project.
Annnnnd Natalia has also posted a doc that goes against VocaTone's statement, too. This situation is the gift that keeps on giving.
 

Alphonse

Passionate Fan
Mar 13, 2021
149
It honestly sounds like this Natalia person is on the world's pettiest revenge campaign. Especially if you're going to hold onto such accusations and only drop them years later for maximum impact.
 

pico

robot enjoyer
Sep 10, 2020
603
VocaTone cannot claim Maghni as a VocaTone project when that identity brings credibility, money, publicity, libraries, institutional relationships, and investor attention, then claim that the project was never truly VocaTone’s when the remaining obligations become inconvenient.
Quotes from the IndieGoGo campaign:
> Maghni AI is an innovative deep-neural-network powered singing synthesizer made by VocaTone Studio and Misbah Studios.
> Maghni AI is a collaboration between Misbah Studios and VocaTone in order to raise the standard of what vocal synthesizers should offer.
> Natalia Shmueli: DIRECTOR, CEO of Misbah Studios
> Austin Grissom: DIRECTOR, CEO of VocaTone Studio
> Creator responsible for this project: Natalia Shmueli doing business as VocaTone Studio.

Are we not missing the forest for the trees? I hate to be pedantic, but is Crescendia not also kind of doing the same thing here? VocaTone says "not my problem" and points the finger at Crescendia. Crescendia says "not my problem" and points the finger at VocaTone. Isn't everyone passing the buck when it suits them?

I think VocaTone is making the case that Natalia was the IndieGoGo's administrator & responsible for the money management, hence technically responsible to fulfill financial obligations. I think Natalia is essentially making the case 'it's more complicated than that and they should step up since we used their brand to garner investment', which may or may not hold up.

It seems the only way anyone will know for sure is if this goes to court. Maybe everyone should put the google doc share links away until seeking legal counsel.

If only business management were as simple as saying "pinky promise" to your homies and wishing on a kickstarter star.

feel like natalia sort of screwed herself here. she goes on record saying she volunteered herself to handle the money because austin outright said he was not qualified to handle the money. doesn't that indicate... that natalia... is responsible for the management of the funds... and the money?

OK so yes, vocatone may have benefitted from how the funds were spent, but if it was your responsibility to manage it... and you mismanaged it and now you cannot fulfill rewards...

just to be 100% absolutely clear, I am not trying to play defense for VocaTone or any of these fools, just kind of awestruck at the literary tragedy playing out before us
 
Last edited:

cafenurse

Still misses Anri Rune
Apr 8, 2018
1,916
24
USA
I had never put much stock into this project because I was never a huge fan of Oliver and when this project was first announced I was distraught it wasn't new vocaloid voicebanks, but my heart does go out to everyone who put money into this and who were really excited for the project. I know many fans in the community love Oliver so this is truly disheartening. I remember being there for Stella, but that incident broke way fewer hearts because she wasn't an already established and beloved character.
 

Bookworm2

Your friendly neighborhood Vocaloid nerd
Holy christ, I miss a single day and some shit goes down.

So. Uh. what the heck is this? What happened? Well, I know what happened, I read the thread, but this is insane. I always kinda doubted that Maghni would come out, but I held out hope because they were likely the only chance of a Big Al update, but with this there's no way. Their reputation has simply been damaged and destroyed to the point that even if by some miracle they manage to pull together a functional engine and release it, I highly doubt that anyone would be willing to buy it.They have shot themselves in the foot.

Re: SynthV usage, yeesh. I kinda get their logic, but if the renders sounded like shit, then don't make a demo yet! If they were so bad that you needed to mess with SynthV and RVC, then clearly the software isn't ready for a release of any kind. Taking a demo also has the problem of of it released in the future, people would probably notice that it sounds very different and comment on that, and this whole thing would have been in vain anyways. It's more work to make a faked demo that will inevitably cause problem than to just wait.

Re: sh*tacon stuff, oh no not freaking again. What is it with VSynth people and underage interests? Can y'all please just not? Good lord.

TL:DR Goddammit I wanted Big Al update, clearly that isn't happening.
 

Users Who Are Viewing This Thread (Users: 0, Guests: 2)